יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

RobustSGPO: בקרת מרחב חיפוש להתפתחות אבזרי סוכנים

RobustSGPO: Search-Space Control for Agent Harness Evolution
RobustSGPO היא שיטה לשיפור אבזרי סוכנים באמצעות משוב מבצעי. היא מאפשרת בקרת מרחב חיפוש ושיפור איכות התוצאות. השיטה נבדקה על 120 משימות והראתה שיפור משמעותי בתוצאות.
תקציר מקורי באנגליתarXiv:2609.09646v1 Announce Type: new Abstract: Semantic-gradient-based prompt optimization (SGPO) improves agent harnesses using execution feedback, but its local update rule leaves the choice of edit scope and operation unresolved. We introduce RobustSGPO, which specifies the requested edit, constructs and checks the patch, and continues search from either the incumbent or retained snapshots. We evaluate permission scheduling, cumulative controls, and task-family transfer in the AgentX brainstorming workflow using 120 tasks, 95 runs, and 7,350 candidate attempts. Periodic $1\to2\to3$ scheduling exceeds fixed maximum permission by 0.28 test-score points. RobustSGPO increases completion on 30 held-out tasks from 60.0% to 80.0% and improves test quality from 3.77 to 4.14 under a 20-million-
קרא במקור המקורי