יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

אימות תקשורת בקודרים פאראלליים: NP-Bench ותכנית תזמון

Verifying Coordination in Parallel Coding Agents: NP-Bench and a Scheduling Planner
אימות תקשורת בקודרים פאראלליים על ידי תכנית תזמון פרואקטיבי. המחברים פיתחו תכנית שמסדרת את העבודה של קודרים פאראלליים כדי למנוע סכסוכים. התכנית נבחנה ב-NP-Bench, סביבת בדיקה שמספקת סצנריות של קודרים פאראלליים. התכנית הצליחה למנוע סכסוכים ב-9 מתוך 9 סצנריות.
תקציר מקורי באנגליתarXiv:2610.07261v1 Announce Type: new Abstract: A team of coding agents can look fine agent by agent yet fail as a team: each passes its own tests while the merged result is broken, and single-agent evaluation never catches it. As teams run several LLM coding agents in parallel on one codebase, the agents collide: two rewrite the same function, one codes against a contract a teammate just changed, and integration fails after the work is done. Most coordination tools react (watch for a conflict, then warn), but at agent speed the warning arrives after the wasted edit. We recast the problem as scheduling: take each work item's declared scope, partition the work into disjoint scopes, and order merges along the producer->consumer graph, all up front. We build this planner into Nerveplane and e
קרא במקור המקורי