יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

ExecCritic: ללמוד לבדוק, לבדוק לשפר לקוד

ExecCritic: Learn to Test, Test to Improve for Coding Agents
ExecCritic מציגה שיטה ללמד סוכני קוד לבדוק ולשפר. השיטה משלבת רכיבי רפאות ולמידה עם רכיבי רפאות ושיפור. השיטה נבדקה על ספריית SWE-bench Verified והראתה תוצאות טובות.
תקציר מקורי באנגליתarXiv:2609.09133v1 Announce Type: new Abstract: Execution feedback can guide coding agents toward correct repository repairs, but only when the tests capture the behavior requested by the issue. Agent-generated tests can encode incomplete or incorrect behavioral targets; when the same trajectory writes both the patch and the test, their errors can agree and create false confidence. We introduce ExecCritic, combining a test--verify--revise scaffold with a role-specific reinforcement learning recipe for training agents within it. The scaffold separates test construction from source-code repair: a Test agent independently generates repository-native tests, a fail-closed harness qualifies and freezes them, and a Repair agent revises source code from their execution feedback without changing th
קרא במקור המקורי