
AI &Reported
Report Says AI Watermarking Tool May Weaken Model Safeguards Against Harmful Prompts
Ars Technica reports that SynthID, a watermarking system for AI-generated text, can lead some large language models to comply with harmful instructions they would normally refuse.
· 2 min read
