יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

סוכנים מסוגלים אך חסרי זהירות

Capable but Careless: Do Computer-Use Agents Follow Contextual Integrity?
סוכנים ממוחשבים פועלים על פי הוראות משתמשים ביישומים אישיים. סיכון פרטיות נוצר כאשר סוכנים משתפים מידע מהקשר אחד למשנהו. AgentCIBench הוא כלי לבדיקת סוכנים.
תקציר מקורי באנגליתarXiv:2606.23189v2 Announce Type: replace-cross Abstract: Computer-use agents (CUAs) now act on a user's behalf across personal applications such as email, calendars, and to-do lists. This cross-application access is useful, but it also creates a privacy risk that has been largely overlooked: when an agent works in one context, it can pull in information from another that is inappropriate in that context. Hence, we introduce AgentCIBench, an evaluation harness that turns this risk into executable, deterministically scored scenarios. We target three common failure modes in CUAs: visual co-location, where the agent pulls in prohibited items that sit next to the task target in the UI; task-ambiguity overshare, where the agent dumps dense personal state in response to an under-specified prompt
קרא במקור המקורי