יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

בטיחות של תקשורת רדומה במערכות סוכנים

Safety of Latent Communication in Multi-Agent Systems
תקשורת רדומה במערכות סוכנים עלולה להיות חשופה להתקפי בטיחות. חוקרים גילו שאפשר לשפר את ההתקפים על ידי אימון קשרים רדומים. ניתן לשפר את הבטיחות של המערכת על ידי שינוי הפרסים של האימון.
תקציר מקורי באנגליתarXiv:2609.39788v1 Announce Type: new Abstract: Latent communication enables multi-agent systems to exchange information directly in internal representation space, reducing the token, computation, and latency overhead of text-based communication. To this end, lightweight trainable links are introduced to map the sender's representations into the receiver's input space. In this work, we show that even benign link training can increase harmful compliance relative to text-based communication while the underlying safety-aligned agents remain unchanged. An attacker can amplify this effect by optimizing the links on harmful query--response pairs or poisoning otherwise benign training data. We further develop a reinforcement-learning attack that rewards harmful compliance alongside benign task pe
קרא במקור המקורי