יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

כשהעולם קובע: התקפות על דגלי עולם סמויים לשליטה בשליטה

When the World Lies: Backdoor Attacks on Latent World Models for Downstream Control
אנשי מחקר גילו שדגלי עולם סמויים יכולים להיות מתקפים על שליטה בשליטה. הם יכולים לשלוט בפעולות של נגנים ולהפיל אותם.
תקציר מקורי באנגליתarXiv:2609.15781v1 Announce Type: cross Abstract: Pretrained world models, learned simulators that encode an observation into a latent state and predict how it evolves under actions, are beginning to be reused as off-the-shelf dynamics backbones for control, like pretrained encoders and language models are reused today. We show that this reuse opens a supply-chain backdoor: an adversary who controls only a released checkpoint can hijack the downstream controller, even though the victim trains and evaluates entirely on clean data and never sees the trigger. The attack encodes no explicit trigger-to-action rule. Instead, the poisoned model routes trigger-bearing observations into a chosen latent region and reshapes the local dynamics there, so that the victim's own optimization (Dreamer-styl
קרא במקור המקורי