LExT: Towards Evaluating Trustworthiness of Natural Language Explanations

ACM Conference on Fairness, Accountability, and Transparency (FAccT 2025), 2025

Krithi Shailya, Shreya Rajpal, Gokul S. Krishnan, and Balaraman Ravindran.

LExT evaluates the trustworthiness of LLM-generated explanations through faithfulness and plausibility across six medical datasets. The work identifies important shortcomings in the reliability of natural-language explanations.