Read original ↗
paperarXivTrust 82 · PrimaryPublished 29d agoLive · 25d ago

From Plausible to Actionable: A Position on LLM Self-Explanations

Large Language Models (LLMs) can generate natural language explanations that rationalize their own decisions, a phenomenon commonly referred to as self-explanations.Such explanations have emerged as a promising direction for explainable artificial intelligence (XAI), particularly for interpreting LLM behavior.However, while self-explanations often appear plausible, whether they faithfully reflect a model's underlying reasoning process remains an open question. In this opinion paper, we argue that self-explanations can be highly plausible, questionably faithful, and yet highly actionable. From

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 47%Northwind AI
  • LinkedLinked via arxiv author · 85%Elize Herrewijnen

    From Plausible to Actionable: A Position on LLM Self-Explanations

  • LinkedLinked via arxiv author · 85%Benedetta Muscato

    From Plausible to Actionable: A Position on LLM Self-Explanations

  • LinkedLinked via arxiv author · 85%Gizem Gezici

    From Plausible to Actionable: A Position on LLM Self-Explanations

  • LinkedLinked via arxiv author · 85%Fosca Giannotti

    From Plausible to Actionable: A Position on LLM Self-Explanations

Related to

authored (incoming)

Related across the graph

Topics