יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

האם סוכנים יכולים לבטוח ביכולותיהם?

Can Agents Trust Their Skills? Uncovering Unsafe Chains of Trust in Skill-Based LLM Agents
TrustProbe הוא כלי לגילוי שרשראות אמון לא בטוחות בסוכנים LLM. הוא מזהה נתיבים פגיעים ב-11 סוכנים פתוחים, ומאשר פגיעות ב-25.1% מניסויי הסוכנים. כלים כמו LangChain ו-OpenAI Agents נפגעים.
תקציר מקורי באנגליתarXiv:2609.39065v1 Announce Type: cross Abstract: LLM agents increasingly rely on installable skills, which are packages of instructions, code, and resources that equip them with task-specific capabilities and, once installed, can be automatically invoked across subsequent user tasks. This creates a chain of trust in which users delegate authority to agents, while agent frameworks admit skill-provided content into the agents' context with insufficient validation, allowing malicious skills to influence agent behavior under that delegated authority. Yet, little is known about whether this trust model adequately constrains untrusted skill content before it reaches security-sensitive operations, or how frequently such trust violations arise in real-world agents. We present TrustProbe, a framew
קרא במקור המקורי