יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

האם ההיסטוריה של האג'נט סופרת כשהקיצור יפגע? אינדקציה חלשה לטראס

Does an Agent's History Tell You When Compaction Will Hurt? A Modest, Bounded Effect on the TRACE Paired-Replay Corpus
ההיסטוריה של האג'נט סופרת כשהקיצור יפגע, אך רק באופן חלש. נמצא כי ההיסטוריה של האג'נט לפני קיצור הקונטקסט חזה נזק לאחר הקיצור, אך רק באופן חלש. המחקר חקר את הקורפוס TRACE של 590 גבולות קיצור של AppWorld, ומצא כי ההיסטוריה של האג'נט לפני קיצור הקונטקסט חזה נזק לאחר הקיצור, אך רק באופן חלש.
תקציר מקורי באנגליתarXiv:2610.08722v1 Announce Type: cross Abstract: Many long-horizon agents compact their context on a global rule, usually a token budget, blind to what the agent was doing. We ask whether the agent's recent behaviour predicts when a compaction will hurt. TRACE's public corpus of 590 harness-triggered AppWorld compaction boundaries replays each boundary from a re-executed prefix state under the pre-compaction context and under the summary, and records the burden of the next actions: calls that error or repeat a call already made. We find that pre-boundary history predicts post-compaction harm only weakly. An internally prespecified contrast by prefix placement is a wide null, and the naive "has-written" label behind it turns out to measure trajectory phase. The best extension-protocol trig
קרא במקור המקורי