יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

האם מודלים ויזואלי-לשוניים מבינים שכנוע ויזואלי?

Do Vision-Language Models Understand Visual Persuasiveness? A Diagnosis via Visual Persuasive Factors
חוקרים בדקו האם מודלים ויזואלי-לשוניים מבינים שכנוע ויזואלי. הם מצאו שהמודלים נוטים לחזור יותר מדי על תמונות כשכנועיות. הם הציגו גם את Visual Persuasive Factors, טקסונומיה לקביעת רמזים ויזואליים.
תקציר מקורי באנגליתarXiv:2511.17036v2 Announce Type: replace Abstract: Visual persuasion uses images to shape cognition, emotion, and behavior, with its effects depending on both visual attributes and semantic context. Despite recent progress, it remains unclear whether Vision-Language Models (VLMs) understand visual persuasiveness. This motivates us to ask: can VLMs assess whether an image persuasively supports an intended message, which visual factors shape this judgment, and do they align with human judgments? Through empirical analyses on image-message pairs where human raters consistently agree on the persuasiveness judgment, we show that VLMs exhibit a recall-oriented bias: they over-predict images as persuasive while achieving high recall. We introduce Visual Persuasive Factors (VPFs), a taxonomy info
קרא במקור המקורי