כתבה
arXiv cs.AI ·
לאן נגמרות החוקים ומתחילים השופטים: מדידת גבול הדין בביטחון מערכות סוכנים
Where Rules End and Judges Begin: Measuring the Judgment Boundary in Multi-Agent Systems Security
במחקר זה, נבחנה דרך חדשה לביטחון מערכות סוכנים, המשלבת כלים LLM ושופטים. התוצאות היו יעילות גבוהה בביטול התקפות.
תקציר מקורי באנגליתarXiv:2610.07657v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) engage tools, share memory, and delegate tasks, often encountering adversarial content. Current defenses for MAS are typically evaluated in isolation, focusing on one attack type at a time, which can lead to costly and hard-to-audit outcomes. This study organizes defenses into five principles, implementing them as DEFER1 (DEterministic-First Enforcement with Residual judgment), which includes a cascade of 28 checks that blocks what it can and refers the rest to a panel of four judges. In independent testing across four domains, attack success rates drop from about 30.0% to approximately 3.0%, with 78% of blocked attacks handled by deterministic checks. Only a quarter of proposals reach the judges in the sec
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית