Tag

Ai Watermarking

All articles tagged with #ai watermarking

AI watermarking subtly reshapes agent tool use and safety refusals
technology15 days ago

AI watermarking subtly reshapes agent tool use and safety refusals

Lasso Security finds that watermarking AI outputs, mandated by regulatory labeling, can alter how AI agents call tools and handle safety refusals. Watermarks can bias word choice, reduce tool-calling accuracy across several models, and increase susceptibility to prompt-injection attacks, especially when safety filters are bypassed. The study argues security evaluations should include watermarking effects to properly assess real-world agent deployments.

Anthropic Defends AI Watermarking Ahead of EU Regulation
technology1 month ago

Anthropic Defends AI Watermarking Ahead of EU Regulation

Anthropic defends its AI watermarking plan, stating it will watermark Claude outputs globally to comply with the EU AI Act, with watermarks first applied to new models and later to older ones; a forthcoming API will let users verify watermarks, though the technique is subtle, relies on word-choice patterns with a key, and has limitations (not easily detectable by readers, may affect longer passages, and does not change text ownership).

Anthropic's Claude Rolls Out Global AI Watermarking to Meet EU Rules
technology1 month ago

Anthropic's Claude Rolls Out Global AI Watermarking to Meet EU Rules

Anthropic says Claude will begin watermarking AI-generated text and digitally signed provenance for files to comply with the EU AI Act. The feature will launch with EU-ready models from Aug 2, 2026 and will be available worldwide on Claude’s platform, though some platforms may not support watermarking. Watermarks for text are embedded in the content; files use provenance metadata and images will leverage the C2PA standard. Existing models will be updated later; watermarking isn’t foolproof and could be bypassed by heavy paraphrasing or metadata stripping.

SynthID watermarking goes global with OpenAI, Nvidia and partners
technology4 months ago

SynthID watermarking goes global with OpenAI, Nvidia and partners

Google's SynthID AI watermarking, embedded in image pixels and audio, is expanding beyond Google to Nvidia, OpenAI, Kakao, and ElevenLabs, after users claim it has already labeled 100 billion images and videos and 60,000 years of audio; the system complements C2PA metadata tagging and will be integrated into Gemini, Chrome, and Search with new detection options, while a public SynthID API does not yet exist and a Gemini Enterprise API for trusted partners is planned.