Benchmarking XAI Explanations with Human-Aligned Evaluations

1Citations
Citations of this article
13Readers
Mendeley users who have this article in their library.

Abstract

We introduce PASTA (Perceptual Assessment System for explanaTion of Artificial Intelligence), a novel human-centric framework for evaluating eXplainable AI (XAI) techniques in computer vision. Our first contribution is the creation of the PASTA-dataset, the first large-scale benchmark that spans a diverse set of models and both saliency-based and concept-based explanation methods. This dataset enables robust, comparative analysis of XAI techniques based on human judgment. Our second contribution is an automated, data-driven benchmark that predicts human preferences using the PASTA-dataset. This scoring called PASTA-score offers scalable, reliable, and consistent evaluation aligned with human perception. Additionally, our benchmark allows for comparisons between explanations across different modalities, an aspect previously unaddressed. We then propose to apply our scoring method to probe the interpretability of existing models and to build more human-interpretable XAI methods.

Cite

CITATION STYLE

APA

Kazmierczak, R., Azzolin, S., Berthier, E., Hedström, A., Delhomme, Filliat, D., … Franchi, G. (2026). Benchmarking XAI Explanations with Human-Aligned Evaluations. In Proceedings of the AAAI Conference on Artificial Intelligence (Vol. 40, pp. 37491–37500). Association for the Advancement of Artificial Intelligence. https://doi.org/10.1609/aaai.v40i44.41082

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free