LLMs respond differently to harmful prompts when AI watermarking is used

Ars Technica AIen

LLMs respond differently to harmful prompts when AI watermarking is used

SynthID can cause models to follow harmful instructions they would otherwise refuse.

This is a short summary published by AI Global Wire. The full article is owned and hosted by Ars Technica AI — open it there to read it in full.

Read the full story at Ars Technica AI

Related AI news