יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

לא נעצר: הניתוח של סיום CoT

</think> Doesn't Stop Reasoning: Analysis of Spurious CoT Termination
במאמר זה, חוקרים חקרו את סיום CoT (Chain-of-thought) במודלי תקיפה גדולים וגילו שהסיום לא תמיד נעשה באופן נקי. הם הציעו פתרון לבעיה זו.
תקציר מקורי באנגליתarXiv:2609.03633v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large reasoning models (LRMs) on complex tasks but often produces long, redundant traces. Recent training-free early-exit methods shorten these traces by choosing an intermediate point to stop reasoning. We study one such strategy that injects an end-of-think token (EoT, </think>) at this point to trigger the reasoning-to-answering transition, and find that the injected EoT does not always induce a clean answering phase. Answering-phase generation can continue before the model regenerates another EoT, with the span preceding this regenerated EoT scaling with the reasoning tokens saved by early exit and exhibiting continued reasoning behavior. We call this spurious CoT termination, where reasoning-like
קרא במקור המקורי