
Same evidence, different judgments: Evidence noncommutative in vision/speech-text conflicts
arXiv:2609.26986v1 Announce Type: new Abstract: For multimodal large language models, when images or speech conflict with accompanying text, measured text reliance can entangle modality preference with evidence position. Earlier studies of text bias often used a fixed evidence order or moved task instructions with the evidence, leaving the contribution of order unclear. In this paper, we use a…
Read original article on cs.AI updates on arXiv.org →