יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

כאשר ראייה גוברת על ידיעה

When Seeing Overrides Knowing: Visual Dominance and Deferral-Based Method for Personalized Safety in VLMs
דגמי שפה-ראייה מותקנים בסביבות בעלות סיכונים גבוהים, שם תגובה סבירה בכלליות עדיין עלולה להיות לא בטוחה עבור משתמש מסוים. MPS-Bench הוא בנך' לבדיקת בטיחות אישית, ו-PRISM הוא מנגנון לזיהוי תחומים שדורשים היסוס.
תקציר מקורי באנגליתarXiv:2609.04281v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed in high-stakes settings, where a response that is reasonable in general may still be unsafe for a particular user whose medical, emotional, or situational context is unknown to the model. We study this problem of personalized safety in multimodal systems and introduce MPS-Bench, a benchmark of 5,181 scenarios from 584 real-world images across 12 high-risk domains, each paired with a hidden user profile. Evaluating eight frontier VLMs, we find that they almost always respond directly (86-99%) rather than seek missing context, and none exceeds 2.6/5 on personalized safety. To understand why these failures arise, we analyze multimodal interactions and identify visual dominance: visual inf
קרא במקור המקורי