כתבה
arXiv cs.AI ·
מדריך מחדש לסוכנים מקודדים
From Verification Failures to Reusable Guidance for Coding Agents
חוקרים פיתחו שיטה ליצירת הנחיה חוזרת לסוכנים מקודדים. הגישה משלבת הגדרות שפה במסגרת K עם קיט של פרוצדורות לבניית תקנים, תיקון הוכחות ובדיקת תוכניות.
תקציר מקורי באנגליתarXiv:2609.39022v2 Announce Type: replace-cross Abstract: Coding agents need to establish that a program satisfies a specification and that the specification captures the requested behavior. We study how expert diagnosis of verification failures can become reusable guidance for this work. Our approach combines executable language definitions in the K framework with a kit of procedures for constructing specifications, repairing proofs, and auditing their adequacy. A human-guided development campaign on HumanEval, a benchmark of 164 Python programming tasks, achieves a 164/164 success rate with the semantics and the kit, measured by final AI audit Pass verdicts after two targeted repairs. To examine whether auditing detects problems that successful proofs leave unresolved, we construct 12 au
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית