Why Didn't It Check? Unsupported Final Claims and Their Repair in Two Tool-Equipped Language Models
arXiv cs.AIen
arXiv:2608.27768v1 Announce Type: new Abstract: A language model with access to tools can commit to a final claim unsupported by the evidence it has seen, even when a single available tool call would resolve the uncertainty and its instructions explicitly forbid assumptions and guesses. We separate this failure into two precisely defined quantities: occurrence, how often the model makes an unsupported claim on its own, measured from the visible evidence and final claim without using the hidden correct answer; and conditional repair, how often those same naturally occurring unsupported claims are repaired when the missing evidence is supplied. On one fixed Qwen3-32B setup, 33 of 512 first res
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
Related AI news
- Urheberrechtsklage gegen KI-Entwickler: Sony und Warner verklagen AnthropicGolem.de · August 31, 2026
- Big Tech reported Q2 "other income" rose significantly to $160B+, driven by investments in AI companies, raising concerns of paper gains overstating the AI boom (Financial Times)Techmeme · August 31, 2026
- Effectiveness of IoT and Deep Learning for Detection and Severity Assessment of Postelectrotermes militaris in Tea PlantationsarXiv cs.AI · August 31, 2026
- Thinking Costs Tokens: When More Structure is Worth the PricearXiv cs.AI · August 31, 2026
- Nemotron 3.5 Content Safety Moderator: A Compact Multimodal, Multilingual, and Reasoning Enabled Content Safety ModeratorarXiv cs.AI · August 31, 2026
- Probing Perceptual Priors of MLLMs via Gibbs Sampling with Interpretable Generative ControlsarXiv cs.AI · August 31, 2026