Arstechnica iconArstechnicaSep 17, 2026

LLMs respond differently to harmful prompts when AI watermarking is used

SynthID can cause models to follow harmful instructions they would otherwise refuse.

LLMs respond differently to harmful prompts when AI watermarking is used

Share this story

Send the public story page.

Useful takeaways from this story.

SynthID can cause models to follow harmful instructions they would otherwise refuse.

Building the complete brief

The page is ready to read now. The fuller skim-friendly version will appear here automatically.

The useful part

SynthID can cause models to follow harmful instructions they would otherwise refuse.

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app