Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools
arXiv cs.AIen
arXiv:2609.05587v1 Announce Type: new Abstract: Existing evaluations of tool-using agents primarily measure whether an agent can successfully complete diverse tasks with tools. These evaluations generally assume that tools return reliable information. However, tool returns in real-world systems can be plausible yet incorrect. We investigate how agents respond to unreliable tool returns by evaluating fourteen LLMs using three tools-web search, LLM sub-agent delegation, and code execution. For each tool, we corrupt its returns and measure whether agents adopt the corrupted content in their final answers. Agents exhibit high levels of overtrust across all three settings: the mean adoption rate
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Agenter
- Företag
Related AI news
- BharatPe launches Gemini-powered AI assistant for merchantsTech in Asia · September 9, 2026
- China's DeepSeek taps CITIC Securities for domestic IPOEconomic Times Tech · September 9, 2026
- Oracle plans HPE networking rollout for AI data centersTech in Asia · September 9, 2026
- Sources: China Securities Regulatory Commission is informally tightening IPO approvals for humanoid startups after a volatile debut by industry leader Unitree (The Information)Techmeme · September 9, 2026
- LG Innotek tackles glass substrate microcrack issue as 2028 production race heats upDIGITIMES · September 9, 2026
- Apple yields to memory suppliers in historic strategy shift, with Kioxia tipped as NAND long-term agreement recipientDIGITIMES · September 9, 2026