Skip to content
Announcement

LLMs respond differently to harmful prompts when AI watermarking is used

SynthID can cause models to follow harmful instructions they would otherwise refuse.

Type
ANNOUNCEMENT
Detection
feed
Confidence
low
Status
Related entities
No tracked entity matched (entity linking by name/alias).

Primary source