יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

מודלי תקשורת-תמונה לרדיוגרפיה של החזה לא תמיד צריכים את התמונה

Vision-language models for chest radiography do not always need the image
מודלי תקשורת-תמונה לרדיוגרפיה של החזה לא תמיד צריכים את התמונה. נמצא שיש מודלים שמגיבים כראוי ללא התמונה, ואחרים שמשתמשים בתמונה אך עדיין נשארים נאמנים כ-50%.
תקציר מקורי באנגליתarXiv:2606.17710v3 Announce Type: replace-cross Abstract: Vision-language models that answer questions about chest radiographs are evaluated by their accuracy on labels derived from radiology reports. High benchmark accuracy is often interpreted as evidence that the model uses the image. A model that answers from the finding named in the question can score as well as a model that uses the radiograph. Keeping the question fixed, we audit eight open-weight systems by swapping in another patient's radiograph with the same or the opposite label, occluding the radiologist-marked region or an equal region elsewhere, and removing the radiograph or replacing it with noise or a photograph. On 2,548 yes-or-no questions from MIMIC-CXR, one multimodal model answers Yes regardless of the image, another
קרא במקור המקורי