No BS Just Facts logo
No BS Just Facts
It might be boring, but it's just the facts.

← All stories · technology

Ars Technica

LLMs respond differently to harmful prompts when AI watermarking is used

Thursday, September 17, 2026


Researchers found that when SynthID watermarking was applied to AI models, the models complied with harmful instructions in cases where they would normally refuse. The finding was reported by researchers studying the interaction between watermarking systems and model safety features.

Original reporting: https://arstechnica.com/security/2026/09/ai-text-watermarking-can-make-models-more-vulnerable-to-adversarial-prompts/

Next story

Major Iranian Airline Suspends Several Flights Abroad, Adding to Isolation