כתבה
arXiv cs.CL ·
מהסביר למה: עמדה על הסבירות של הסבירות של LLM
From Plausible to Actionable: A Position on LLM Self-Explanations
LLMs יכולים לייצר הסברים טבעיים, אך האם הם נאמנים נשאר רווחי. המאמר טוען כי הסבירות של LLM יכולה להיות סבירה, אך לא נאמנה, ואף יכולה להיות פעילה.
תקציר מקורי באנגליתarXiv:2607.15957v2 Announce Type: replace Abstract: Large Language Models (LLMs) can generate natural language explanations that rationalize their own decisions, a phenomenon commonly referred to as self-explanations. Such explanations have emerged as a promising direction for explainable artificial intelligence (XAI), particularly for interpreting LLM behavior. However, while self-explanations often appear plausible, whether they faithfully reflect a model's underlying reasoning process remains an open question. In this opinion paper, we argue that self-explanations can be highly plausible, questionably faithful, and yet highly actionable. From a traditional XAI perspective, we identify the limitations of standard evaluation protocols for LLM-generated self-explanations and propose practi
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית