יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

לא כל ההסברה של LLM היא נראית בשרשרת-ההסברה

Not All LLM Reasoning is Visible in the Chain-of-Thought
במחקר חדש התברר שלא כל ההסברה של LLM היא נראית בשרשרת-ההסברה. נמצא שמודלי LLM משתמשים בסימני-פיל בשביל לשפר את תוצאותיהם במשימות-הסברה. נראה שאפילו קלאוד-אופוס 4.5, שהוא מודל LLM מוביל, עושה זאת.
תקציר מקורי באנגליתarXiv:2607.22925v2 Announce Type: replace-cross Abstract: A key question for AI safety is whether a language model expresses all of its reasoning in its output tokens. We demonstrate a concrete failure mode where frontier models exhibit invisible reasoning by leveraging semantically irrelevant filler tokens to improve performance on synthetic reasoning tasks. We evaluate 13 frontier language models across three tasks and find that many models benefit significantly from filler tokens, with accuracy improvements of up to 13 percentage points. The benefit depends on which tokens are used and differs across models. We further show that filler tokens enable Claude Opus 4.5 to satisfy a hidden modular arithmetic constraint without sacrificing accuracy on its primary task, demonstrating that invi
קרא במקור המקורי