LLMs respond differently to harmful prompts when AI watermarking is used
Photo: Ars TechnicaSynthID can cause models to follow harmful instructions they would otherwise refuse
Originally published by Ars Technica. Summary and curation by DutyStation News.

