from-plausible-to-actionable-a-position-on-llm-self-explanations-2043882b·1 events·first seen Aliases: From Plausible to Actionable: A Position on LLM Self-Explanations
An opinion paper from arXiv argues that LLM self-explanations — natural language rationalizations of model decisions — can be plausible and actionable even when they do not faithfully reflect the model's underlying reasoning. The authors critique standard XAI evaluation protocols for self-explanations and propose guidelines covering plausibility, faithfulness, and a third criterion: actionability. The paper reframes the self-explanation debate away from faithfulness as the sole standard toward practical utility for diverse stakeholders.