Pith. sign in

REVIEW 12 cited by

LLMs for Explainable AI: A Comprehensive Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.00125 v1 pith:CVF2WCX5 submitted 2025-03-31 cs.AI cs.CL

LLMs for Explainable AI: A Comprehensive Survey

classification cs.AI cs.CL
keywords modelsllmsusersexplainablemodelapproachescomprehensivehuman
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Large Language Models (LLMs) offer a promising approach to enhancing Explainable AI (XAI) by transforming complex machine learning outputs into easy-to-understand narratives, making model predictions more accessible to users, and helping bridge the gap between sophisticated model behavior and human interpretability. AI models, such as state-of-the-art neural networks and deep learning models, are often seen as "black boxes" due to a lack of transparency. As users cannot fully understand how the models reach conclusions, users have difficulty trusting decisions from AI models, which leads to less effective decision-making processes, reduced accountabilities, and unclear potential biases. A challenge arises in developing explainable AI (XAI) models to gain users' trust and provide insights into how models generate their outputs. With the development of Large Language Models, we want to explore the possibilities of using human language-based models, LLMs, for model explainabilities. This survey provides a comprehensive overview of existing approaches regarding LLMs for XAI, and evaluation techniques for LLM-generated explanation, discusses the corresponding challenges and limitations, and examines real-world applications. Finally, we discuss future directions by emphasizing the need for more interpretable, automated, user-centric, and multidisciplinary approaches for XAI via LLMs.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 12 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. XGRAG: A Graph-Native Framework for Explaining KG-based Retrieval-Augmented Generation

    cs.AI 2026-04 unverdicted novelty 7.0

    XGRAG uses graph perturbations to quantify component contributions in GraphRAG and achieves 14.81% better explanation quality than text-based baselines on QA datasets, with correlations to graph centrality.

  2. Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs

    cs.AI 2026-04 unverdicted novelty 6.0

    COMPASS formalizes prompt engineering as a POMDP-based cognitive decision process for self-adaptive generation of task plan explanations via LLMs.

  3. M2-PALE: A Framework for Explaining Multi-Agent MCTS--Minimax Hybrids via Process Mining and LLMs

    cs.AI 2026-04 unverdicted novelty 6.0

    M2-PALE extracts process models from multi-agent MCTS-Minimax execution traces using Alpha Miner, iDHM and Inductive Miner, then uses LLMs to generate causal explanations, shown in a small checkers setting.

  4. Quantifying Trust: Financial Risk Management for Trustworthy AI Agents

    cs.AI 2026-04 unverdicted novelty 6.0

    The paper introduces the Agentic Risk Standard (ARS) as a payment settlement framework that delivers predefined compensation for AI agent execution failures, misalignment, or unintended outcomes.

  5. When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making

    cs.AI 2026-02 unverdicted novelty 6.0

    Adversarial explanation attacks preserve nearly all human trust in wrong AI outputs by using persuasive framing, shown in a study varying reasoning, evidence, style, and format with over 200 participants.

  6. SMT-AD: a scalable quantum-inspired anomaly detection approach

    cs.LG 2026-04 unverdicted novelty 5.0

    SMT-AD detects anomalies via superposed multiresolution bond-dimension-1 MPOs with Fourier embedding, claiming competitive baseline performance and linear parameter scaling.

  7. SMT-AD: a scalable quantum-inspired anomaly detection approach

    cs.LG 2026-04 unverdicted novelty 5.0

    SMT-AD applies superposition of bond-dimension-1 matrix product operators with multiresolution Fourier embedding to achieve competitive anomaly detection on standard datasets with linear parameter growth.

  8. AnTenA: Actionable and Explainable Tensor Analysis System with Large Language Models

    cs.CL 2026-06 unverdicted novelty 4.0

    AnTenA uses task-agnostic and task-specific LLM prompts to explain co-clustered patterns from tensor decomposition and evaluates them on forward and backward inference tasks.

  9. Explainable AI for Next-Generation Wireless Physical Layer: Basics, State-of-the-Art, and Open Challenges

    eess.SP 2026-06 unverdicted novelty 4.0

    A survey formalizing responsibility-oriented goals for wireless XAI, developing a taxonomy of explainability approaches, reviewing PHY layer applications, and discussing open challenges including performance tradeoffs...

  10. ExAI5G: A Logic-Based Explainable AI Framework for Intrusion Detection in 5G Networks

    cs.CR 2026-04 unverdicted novelty 4.0

    ExAI5G combines Transformer-based intrusion detection with surrogate decision trees and LLM-evaluated explanations to deliver 99.9% accuracy and 16 high-fidelity logical rules on 5G IoT traffic while preserving performance.

  11. Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

    cs.CR 2026-04 unverdicted novelty 4.0

    Encoder-based LLMs detect SDN intrusions with decisions driven by meaningful traffic behaviors, as validated by attribution analysis aligning with established intrusion principles.

  12. Bayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering

    cs.AI 2026-07 unverdicted novelty 3.0

    The study applies Bayesian uncertainty propagation to agentic RAG pipelines on StrategyQA and HotpotQA, reporting better discrimination on HotpotQA than on StrategyQA using standard calibration and selective-predictio...