LLMs respond differently to harmful prompts when AI watermarking is used
Ars Technica AIen

SynthID can cause models to follow harmful instructions they would otherwise refuse.
This is a short summary published by AI Global Wire. The full article is owned and hosted by Ars Technica AI — open it there to read it in full.
Read the full story at Ars Technica AIRelated AI news
- NYT court filing: ChatGPT's head wrote that publishers face an "existential threat" and a Microsoft executive called AI training "an astonishing theft" (Financial Times)Techmeme · September 17, 2026
- Gemini 3.8 Live Transforms Conversational AIAI Business · September 17, 2026
- The AI Superintelligence SlowdownThe Verge · September 17, 2026
- OpenAI reportedly closes in on solving the Hodge conjecture, its second Millennium Prize ProblemThe Decoder · September 17, 2026
- Tekoälyn piti auttaa – Tietoturvaosaajat paljastavat karun totuudenTivi · September 17, 2026
- Claude Code relaunches Projects to manage multiple AI agents in the cloudThe Verge · September 17, 2026