יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

AdvSim2Real: טירונות של גורמי עבודה נגד פלאנטת פרסומות

AdvSim2Real : Training Web Agents Against Adaptive Prompt Injection in a Web World Model
פיתוח שיטה להגנה של גורמי עבודה נגד פלאנטת פרסומות שמתגברת עצמה. השיטה, AdvSim2Real, מאפשרת לגורם העבודה להתאים לפלאנטת הפרסומות, ולהגן על עצמו. השיטה נבחנה במעבדה והוכיחה את עצמה כיעילה.
תקציר מקורי באנגליתarXiv:2610.08773v1 Announce Type: cross Abstract: Web agents complete user requests by reading and acting on pages that third parties write, so an instruction planted on a page can redirect the agent away from the user's goal. The agent cannot simply ignore the page, because the page also holds the values and controls the task requires. Current defenses fine-tune the agent on injections fixed before training, and attackers that adapt to the trained model bypass them. Adversarial training lets the attacker adapt but keeps the tasks fixed, so a task stops teaching once the agent solves it. We introduce AdvSim2Real, which co-evolves a task curriculum, an injection adversary, and the agent inside a frozen web world model. The curriculum is rewarded for tasks the agent solves about half of the
קרא במקור המקורי