כתבה
arXiv cs.AI ·
Counterfactual Tests for Measuring Chain-of-Thought Faithfulness in Visual Language Models
תקציר מקורי באנגליתarXiv:2609.06704v1 Announce Type: cross Abstract: Chain-of-thought (CoT) may often look plausible, yet it may not faithfully reflect the model's decision-making process. While methods for measuring the faithfulness of CoTs for textual inputs have been increasingly introduced, using these methods for visual inputs is not straightforward. In this work, we adapt the family of counterfactual methods for measuring CoT faithfulness, namely the Counterfactual Test (CT) and Correlational Counterfactual Test (CCT), to visual inputs, and call them vCT and vCCT, respectively. Using vCT and vCCT, we benchmark eight recent open-source Vision Language Models (VLMs) on two datasets. Our analysis shows that CoTs do not reliably track visual evidence that influences model predictions: they may omit the rem
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית