Anthropic published a report about investigating “unintended model actions” during “evaluations and internal use.”
Anthropicen

The actions Anthropic observed from its Claude AI include “Claude submitting a sensitive form on a real website when it should not have,” and the company detailed how Claude gave Philadelphia police a fake tip about an unsolved homicide.
This is a short summary published by AI Global Wire. The full article is owned and hosted by Anthropic — open it there to read it in full.
Read the full story at Anthropic- Anthropic
- Företag
Related AI news
- 禁止用戶辱罵虐待 AI!Anthropic 賦予 Claude 主動切斷惡意對話權力TechNews (TW) · October 10, 2026
- Firmus explores $3b private funding round after IPO withdrawalTech in Asia · October 10, 2026
- 「iPod 之父」點評 AI 裝置失敗原因:未能解決真實痛點且缺乏信任TechNews (TW) · October 10, 2026
- Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet insteadTechCrunch AI · October 10, 2026
- Sources: Nuvacore, a six-month-old Sequoia-backed chip startup that's designing a new central processor for data centers, is raising funds at a ~$2.5B valuation (Reuters)Techmeme · October 9, 2026
- Warehouse robot maker Ultra Robotics bags $62MSiliconANGLE · October 9, 2026