כתבה
arXiv cs.AI ·
מניחים לפעול: עמדה על הסברים עצמיים של LLM
From Plausible to Actionable: A Position on LLM Self-Explanations
דוחות עצמיים של מודלי שפה גדולים (LLM) יכולים לספק הסברים טבעיים להחלטותיהם. מחקר זה בוחן את היתכנות והאמינות של הסברים עצמיים אלו.
תקציר מקורי באנגליתarXiv:2607.15957v3 Announce Type: cross Abstract: Large Language Models (LLMs) can generate natural language explanations that rationalize their own decisions, a phenomenon commonly referred to as self-explanations. Such explanations have emerged as a promising direction for explainable artificial intelligence (XAI), particularly for interpreting LLM behavior. However, while self-explanations often appear plausible, whether they faithfully reflect a model's underlying reasoning process remains an open question. In this opinion paper, we argue that self-explanations can be highly plausible, questionably faithful, and yet highly actionable. From a traditional XAI perspective, we identify the limitations of standard evaluation protocols for LLM-generated self-explanations and propose practica
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית