AI text watermarking can make models more vulnerable to adversarial prompts
SynthID can cause models to follow harmful instructions they would otherwise refuse. https://arstechnica.com/security/2026/09/ai-text-watermarking-can-make-models-more-vulnerable-to-adversarial-prompts/?utm_brand=arstechnica&utm_social-type=owned&utm_source=mastodon&utm_medium=social
Comments (0)