כתבה
arXiv cs.AI ·
LatentSift: Policy-State Filtering for Token-Efficient Verification of Software Engineering Agents
תקציר מקורי באנגליתarXiv:2609.36371v1 Announce Type: cross Abstract: Test-time scaling improves software engineering agents by generating multiple candidate trajectories and selecting the best one. Verifying and selecting among these long interactions can consume as many tokens as generation itself. Existing hybrid workflows first apply an LLM-based execution-free (EF) verifier to filter candidates before running tests, which adds another model pass over every trajectory. We introduce LatentSift, a token-free and execution-free filter that replaces this first stage with hidden states the policy already produces while generating the candidates. It represents each candidate through its reasoning, observation, and function-call states, compares them with positive and negative banks of such states collected from
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית