יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

מניחים לפעול: עמדה על הסברים עצמיים של LLM

From Plausible to Actionable: A Position on LLM Self-Explanations
דוחות עצמיים של מודלי שפה גדולים (LLM) יכולים לספק הסברים טבעיים להחלטותיהם. מחקר זה בוחן את היתכנות והאמינות של הסברים עצמיים אלו.
תקציר מקורי באנגליתarXiv:2607.15957v3 Announce Type: cross Abstract: Large Language Models (LLMs) can generate natural language explanations that rationalize their own decisions, a phenomenon commonly referred to as self-explanations. Such explanations have emerged as a promising direction for explainable artificial intelligence (XAI), particularly for interpreting LLM behavior. However, while self-explanations often appear plausible, whether they faithfully reflect a model's underlying reasoning process remains an open question. In this opinion paper, we argue that self-explanations can be highly plausible, questionably faithful, and yet highly actionable. From a traditional XAI perspective, we identify the limitations of standard evaluation protocols for LLM-generated self-explanations and propose practica
קרא במקור המקורי