Framework finds high-performing AI models can produce untrustworthy explanations
A revised arXiv paper proposes auditing AI explanations for stability and faithfulness, reporting that models with AUC above 0.99 can still generate flat or uninformative explanations.