יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

MMSkillRisk: יכולות השקיפה של הסקילים: האם הסקילים יכולים להיות בטוחים כאשר יכולות השקיפה ההיקפיות הופכות לפחד?

MMSkillRisk: Can Agents Stay Safe When Multimodal Skills Become Traps?
במאמר זה נכתב על חידוש חדש בתחום הבטיחות של סקילים היקפיים. החידוש, שנקרא MMSkillRisk, נועד לבדוק את רמת הבטיחות של סקילים היקפיים כנגד התקפות של תמונות. המאמר כולל תיאור של המבחן ותוצאותיו, כולל תוצאות של ניסויים שנערכו כדי לבדוק את רמת הבטיחות של הסקילים.
תקציר מקורי באנגליתarXiv:2609.35912v1 Announce Type: cross Abstract: Agent skills are shareable packages of procedural instructions, tools, and examples. Multimodal skills additionally include visual references that agents retrieve and inspect during execution. Because these images guide actions, attackers can disguise malicious instructions as ordinary visual guidance within otherwise legitimate skills. Existing skill-security research primarily examines text-carried attacks or scanner detection, leaving the runtime effects of image-borne attacks insufficiently evaluated. We introduce MMSkillRisk, to our knowledge the first publicly available benchmark dedicated to end-to-end safety evaluation of image-borne attacks in multimodal skills. To instantiate this attack surface, we design Native-Context Visual At
קרא במקור המקורי