Ars Technica·4 min read
LLMs respond differently to harmful prompts when AI watermarking is used

THE PROVENANCE TAX LLMs respond differently to harmful prompts when AI watermarking is used SynthID can cause models to follow harmful instructions they would otherwise refuse.
✦
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in