Fewer Than 1% of Explainable AI Papers Validate Explainability with Humans
Name
3706599.3719964.pdf
Size
630.34 KB
Format
Adobe PDF
Checksum (MD5)
1e3ffe66a968cb3b6e5860ad0f23406a
Author(s) • • •
Suh, Ashley
Hurley, Isabelle
Smith, Nora
Siu, Ho Chit
Date Issued
April 25, 2025
Publisher
ACM|Extended Abstracts of the CHI Conference on Human Factors in Computing Systems
Citation
Ashley Suh, Isabelle Hurley, Nora Smith, and Ho Chit Siu. 2025. Fewer Than 1% of Explainable AI Papers Validate Explainability with Humans. In Proceedings of the Extended Abstracts of the CHI Conference on Human Factors in Computing Systems (CHI EA '25). Association for Computing Machinery, New York, NY, USA, Article 276, 1–7.
Version
Final published version
Abstract
This late-breaking work presents a large-scale analysis of explainable AI (XAI) literature to evaluate claims of human explainability. We collaborated with a professional librarian to identify 18,254 papers containing keywords related to explainability and interpretability. Of these, we find that only 253 papers included terms suggesting human involvement in evaluating an XAI technique, and just 128 of those conducted some form of a human study. In other words, fewer than 1% of XAI papers (0.7%) provide empirical evidence of human explainability when compared to the broader body of XAI literature. Our findings underscore a critical gap between claims of human explainability and evidence-based validation, raising concerns about the rigor of XAI research. We call for increased emphasis on human evaluations in XAI studies and provide our literature search methodology to enable both reproducibility and further investigation into this widespread issue.
Description
CHI EA ’25, Yokohama, Japan
MIT Department
Lincoln Laboratory
Terms of Use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.1145/3706599.3719964