SOURCE-LINKED INTELLIGENCE
LLMs respond differently to harmful prompts when AI watermarking is used
SynthID can cause models to follow harmful instructions they would otherwise refuse.
Read original source ↗ Open in workspace
- recordType
- article
- region
- Global
Evidence & attribution
- Ars Technica · 2026-09-17T18:33:13.000Z
First collected: 2026-09-19T20:26:32.566Z. This is not the publication date.