יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

בירור תקציבי של פרויקציית תקשורת עם רצפי עוזרים שלא נרץ

Evaluating Budgeted Context Projection with Unexecuted Companion Runs
במאמר זה, נבדוק כיצד רצפי עוזרים שלא נרץ יוצרים סכסוך בין פרויקציית תקשורת תקציבית לבין השלמת משימה.
תקציר מקורי באנגליתarXiv:2609.31381v2 Announce Type: replace Abstract: Context projection can shorten individual requests while changing whether an agent finishes within its budget. We examine how sequential evaluation obscures this trade-off when a capped first continuation prevents its companion from running. In a recorded ReVerPi source-reading campaign, 15 pairs with two final answers yield 12 historically scored successes per arm. Retaining all 27 intervention boundaries distinguishes observed failures from ten unexecuted companions and bounds projected-minus-full success between $-9$ and $+1$ tasks. Under the archived scoring contract, a frozen projection selector has a success difference from full context of $[-3,0]$; outside four fitting tasks, it is $[-4,-1]$ across 23 boundaries. Excluding one task
קרא במקור המקורי