יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

משפחה נכונה, מיומנות שגויה: בחינת סיכון חשיפה בחיפוש מיומנויות

Right Family, Wrong Skill: Evaluating Risk Exposure in Agent Skill Retrieval
במאמר זה נחקר סיכון חשיפה בחיפוש מיומנויות של סוכנים. נבחן כיצד סוכנים יכולים להציג מיומנויות שמתאימות לנושא, אך עשויות להתנגש עם דרישות המשאבים, התהליך או התוצאות.
תקציר מקורי באנגליתarXiv:2606.10388v4 Announce Type: replace-cross Abstract: A skill can match a task's topic while conflicting with its resource, procedure, or output requirements. We study this as same-capability risk-exposure retrieval and introduce SameCapRisk-Bench: 890 units and 1,314 query cases across five mechanisms and twelve conflict types. Each unit pairs a skill that meets a query requirement with a same-capability skill that violates it, with evidence for that distinction. The evaluation tests source-task requirements and two kinds of paired queries that reverse which skill fits: changing the requested evidence role or exact output interface while holding the skills fixed. To capture both retrieval success and conflicting exposure, Recall tracks helpful hits, harmful sibling rate (HSR) tracks e
קרא במקור המקורי