禁止用戶辱罵虐待 AI!Anthropic 賦予 Claude 主動切斷惡意對話權力
TechNews (TW)zh-TW

為禁止用戶沒有明顯目的卻反覆辱罵虐待 AI 模型,Anthropic 近日更新使用政策,明令禁止用戶對 AI […]
This is a short summary published by AI Global Wire. The full article is owned and hosted by TechNews (TW) — open it there to read it in full.
Read the full story at TechNews (TW)- Anthropic
Related AI news
- Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet insteadTechCrunch AI · October 10, 2026
- Sources: top execs at Anthropic, OpenAI, and others are gaming out scenarios for a public and political revolt following a catastrophic AI event (Maria Curi/Axios)Techmeme · October 9, 2026
- Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicideThe Verge AI · October 9, 2026
- An Anthropic AI model sent a false homicide tip to Philadelphia policeTechCrunch AI · October 9, 2026
- Anthropic's Claude can now orchestrate up to 1,000 AI agents in parallel through dynamic workflowsThe Decoder · October 9, 2026
- Anthropic launches a free AI scanner for open-source projectsThe Decoder · October 9, 2026