Ars TechnicaCenter·
LLMs respond differently to harmful prompts when AI watermarking is used
SynthID can cause models to follow harmful instructions they would otherwise refuse.
Read at Ars Technica →
SynthID can cause models to follow harmful instructions they would otherwise refuse.

<iframe src="https://statesidedaily.com/embed/cluster/60e1a1fd73976bfe748a26216bc5c91594968906" width="100%" height="320" style="border:0;max-width:680px" loading="lazy" title="LLMs respond differently to harmful prompts when AI watermarking is used"></iframe>