כתבה
arXiv cs.CL ·
קודם לא רוטה: הסבר לפער הטבלה-תרשים באימות טענות וירטואליות
Encoded but Not Routed: Explaining the Table-Chart Gap in Scientific Claim Verification
חוקרים גילו ש-LLMs רב-תכליתיים נכשלים להשתמש במידע מתרשימים כדי ליצור תחזיות, אף על פי שהם קודדים אותו.
תקציר מקורי באנגליתarXiv:2606.01679v2 Announce Type: replace Abstract: Multimodal LLMs are increasingly used to assist scientific peer review, where a core requirement is verifying whether claims in a paper are supported by its evidence. Prior work has shown that models perform substantially better at this task when the evidence is a table than when it is a chart of the same underlying data. This raises the question of whether models fail to extract information from charts, or do they extract it but fail to use it when forming their prediction? We study this question through layer-wise linear probing and attention analysis on three open-weight VLMs over table and chart evidence, representing the same underlying data. We find consistent evidence for the latter. Chart information is encoded in the models' inte
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית