יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

גבולות שיפוט במערכות רב-סוכנים

Where Rules End and Judges Begin: Measuring the Judgment Boundary in Multi-Agent Systems Security
חוקרים פיתחו שיטה למדידת גבולות שיפוט במערכות רב-סוכנים. השיטה, DEFER1, כוללת 28 בדיקות שחוסמות התקפות ומפנות את השאר לפאנל שופטים. ניסויים עצמאיים הראו ירידה בשיעור ההתקפות המוצלחות מ-30% ל-3%
תקציר מקורי באנגליתarXiv:2610.07657v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) engage tools, share memory, and delegate tasks, often encountering adversarial content. Current defenses for MAS are typically evaluated in isolation, focusing on one attack type at a time, which can lead to costly and hard-to-audit outcomes. This study organizes defenses into five principles, implementing them as DEFER1 (DEterministic-First Enforcement with Residual judgment), which includes a cascade of 28 checks that blocks what it can and refers the rest to a panel of four judges. In independent testing across four domains, attack success rates drop from about 30.0% to approximately 3.0%, with 78% of blocked attacks handled by deterministic checks. Only a quarter of proposals reach the judges in the s
קרא במקור המקורי