Interestana
Home/News/Anthropic's Claude AI to Implement Invisible Watermarks for EU AI Act Compliance
Ars Technica••3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Anthropic's Claude AI to Implement Invisible Watermarks for EU AI Act Compliance

Anthropic's Claude AI to Implement Invisible Watermarks for EU AI Act Compliance

Anthropic, the artificial intelligence company behind the Claude family of large language models, has announced its intention to implement machine-readable watermarks on all content processed by its AI systems. This significant move is primarily driven by the impending requirements of the European Union's AI Act. The EU AI Act, a landmark piece of legislation aimed at regulating artificial intelligence, mandates that providers of AI systems must watermark AI-generated or manipulated audio, image, text, and video outputs. This regulation specifically targets AI models released after August 2, 2024, and provides a grace period until December 2026 for existing models to be brought into compliance. Anthropic has confirmed that its new models, regardless of their geographical release location, will incorporate these watermarks for AI-generated content from their inception. For text outputs, these watermarks will be "embedded" within the content, rendering them invisible to the end-user. For other file types, such as images or audio, Anthropic plans to include "digitally signed provenance metadata where supported," providing a verifiable trail of the content's origin. A notable aspect of Anthropic's strategy is its adoption of a comprehensive approach, which it describes as "nuke it from orbit." This means the company intends to apply watermarks to all processed content where technically feasible, even in scenarios that the EU AI Act explicitly exempts. The EU guidance, for instance, carves out exceptions for AI systems performing "an assistive function for standard editing" – such as grammar correction – or when the AI does not "substantially alter" the user's original text or its meaning. However, by implementing watermarking at the model level, Anthropic acknowledges that its Claude models may inadvertently watermark content that has undergone only minor modifications, like a simple comma correction, which the law was designed to leave unmarked. The true efficacy and scope of these watermarks will not be fully ascertainable until Anthropic releases a dedicated detection tool. This tool will allow for independent testing and verification of the watermarking system. Anthropic has stated its commitment to eventually sharing technical details regarding watermark detection, a move that aligns with the EU law's requirement for providers to offer technical support to users and authorities. This proactive stance by Anthropic positions the company to meet evolving global AI regulations and enhance transparency in the use of its AI technologies, though the practical implications for content creators and users will become clearer with the release of the detection mechanisms.

Original source — read the full reporting at the publisher:

Read on Ars Technica

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next