Announcement
LLMs respond differently to harmful prompts when AI watermarking is used
SynthID can cause models to follow harmful instructions they would otherwise refuse.
- Type
- ANNOUNCEMENT
- Detection
- feed
- Confidence
- low
- Status
Related entities
No tracked entity matched (entity linking by name/alias).