כתבה
arXiv cs.CL ·
When Reasoning Goes Astray: Attention Dynamics of Uncontrolled Reasoning
תקציר מקורי באנגליתarXiv:2609.38817v1 Announce Type: cross Abstract: Large reasoning models (LRMs) improve performance on complex tasks through extended reasoning, yet the same process can degenerate into redundant verification and persistent generation loops. Such uncontrolled reasoning increases inference cost and creates risks of resource exhaustion and service degradation. However, existing mitigations largely truncate long outputs or react to surface repetition, and thus fail to distinguish normal thinking from uncontrolled reasoning or explain how benign reasoning degenerates into harmful behavior. In this paper, we operationalize LRM generation as four states and further introduce Reasoning-state Analysis via Dynamic Attention Responses (RADAR), which identifies the current reasoning state in real tim
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית