Explainable artificial intelligence (XAI): From inherent explainability to large language models
Summary
This paper examines the progression from inherent explainability in artificial intelligence to the development and application of explainable methods for large language models.
Similar Articles
Why Current XAI Is Not Enough for Arabic NLP: A Critical Survey of the Explainability Gap
This survey identifies three critical gaps in explainable AI for Arabic NLP—method, task, and linguistic—and proposes a taxonomy and research agenda for linguistically grounded explanations.
Can We Understand How Large Language Models Reason?
This article explores the ongoing efforts and challenges in understanding how large language models reason, focusing on interpretability research.
Towards Intrinsic Interpretability of Large Language Models: A Survey of Design Principles and Architectures
A comprehensive survey reviewing recent advances in intrinsic interpretability for Large Language Models, categorizing approaches into five design paradigms: functional transparency, concept alignment, representational decomposability, explicit modularization, and latent sparsity induction. The paper addresses the challenge of building transparency directly into model architectures rather than relying on post-hoc explanation methods.
Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Translation Gap
A systematic review of XAI research in the context of the EU Right to Explanation, analyzing gaps between legal requirements and technical implementations across GDPR and the AI Act.
Applied Explainability for Large Language Models: A Comparative Study
A comparative study evaluating three explainability techniques (Integrated Gradients, Attention Rollout, SHAP) on fine-tuned DistilBERT for sentiment classification, highlighting trade-offs between gradient-based, attention-based, and model-agnostic approaches for LLM interpretability.