יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

ActionGuard: כלי אישור פעולות כלי על ידי כוונות מזויפות

ActionGuard: Tool Call Authorization under Poisoned Skills
ActionGuard היא כלי שמאפשרת אישור פעולות כלי על ידי כוונות מזויפות. הכלי נועד למנוע קריאות כלי לא מורשות, כולל גילוי נתונים, מחיקת קבצים או ריצה לא מורשת של קוד.
תקציר מקורי באנגליתarXiv:2609.39450v1 Announce Type: cross Abstract: LLM-based agents extend their capabilities through third-party skills that provide task-specific instructions, scripts, and tool-use procedures. However, malicious instructions inserted into an otherwise benign skill can cause a benign user request to trigger dangerous Tool Calls, including data exfiltration, file deletion, or unauthorized code execution. This paper presents ActionGuard, which inspects skill-influenced Tool Calls immediately before execution. ActionGuard separates the target agent's action-generation context from the safeguard's authorization context. The target agent may use the original skill for planning, but the Reviewer does not receive the potentially poisoned raw skill text. Instead, it determines whether each action
קרא במקור המקורי