Same evidence, different judgments: Evidence noncommutative in vision/speech-text conflicts
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.26986v1 Announce Type: new Abstract: For multimodal large language models, when images or speech conflict with accompanying text, measured text reliance can entangle modality preference with evidence position. Earlier studies of text bias often used a fixed evidence order or moved task instructions with the evidence, leaving the contribution of order unclear. In this paper, we use a paired comparison that keeps the instructions and evidence content fixed and swaps only the positions of the two sources to quantify this potential influence. Across vision and speech models, placing an image or recording after conflicting text consistently shifts answers toward its content. We also re
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Bild
Related AI news
- Experts say that air-gapping AI could prevent events like the Hugging Face hack, but would undermine the value of evaluations and slow research to a crawl (Robert Hart/The Verge)Techmeme · September 25, 2026
- Google’s first Project Suncatcher AI satellite set to blast off into orbit next weekSiliconANGLE · September 25, 2026
- Singapore finance firms aim to train 80,000 workers in AITech in Asia · September 25, 2026
- US AI video startup Higgsfield eyes $1b annual salesTech in Asia · September 25, 2026
- Researchers link more cyberattacks to OpenAI agent swarmSiliconANGLE · September 24, 2026
- Meta unveils mobile app Horizon Create and web app Horizon Studio for building games with AI prompts; the games will run on Facebook, Instagram, and Horizon (Jay Peters/The Verge)Techmeme · September 24, 2026