Anthropic Shares Details on Claude’s New Watermarks to Comply with EU AI Act
Anthropic’s Watermarking Efforts: A Comprehensive Explanation
Anthropic, the developer of the popular chatbot Claude, has recently shared more details about its efforts to implement watermarks in its generated text. The move is aimed at complying with the EU AI Act’s Transparency Code, which requires AI companies to use systems that make it possible to identify AI-generated content.
The debate surrounding the watermarks has been ongoing, with some users expressing concerns that the watermarks could be removed or hidden. However, Anthropic has assured its users that the watermarks will not impact the quality of Claude’s output and will be indistinguishable from unwatermarked responses to human readers.
How Watermarking Works
Anthropic uses the SynthID-Text approach, which was outlined by the Google DeepMind team in 2024. This approach involves creating a pattern in the responses that is undetectable to the reader but can be detected by those with the necessary key. The company plans to release a watermark detection API to facilitate this process.
The watermarks are designed to be resistant to light editing, but a complete rewrite of the text could potentially remove the watermark. However, Anthropic notes that this would also render the text no longer AI-generated, as the watermark is a fundamental aspect of the AI’s output.
Watermarking in Code
Anthropic also addressed the issue of watermarking in code, stating that it should have a negligible effect on the actual code produced. However, in areas where there is an arbitrary choice between particular words or terms within the code, the watermark can be used, such as in comments within the code.
The company emphasizes that other major model developers have signed the same Code of Practice and will be implementing their own watermarks, making it unlikely that users will be able to easily remove or hide the watermarks.
Impact on Users
Some users have expressed concerns that the watermarks could be used to identify and flag AI-generated content as such. However, Anthropic assures its users that the watermarks are designed to be transparent and will not impact the quality of Claude’s output.
The company’s efforts to implement watermarks are aimed at complying with the EU AI Act’s Transparency Code and promoting transparency in AI-generated content.
By releasing more details about its watermarking efforts, Anthropic aims to address user concerns and provide clarity on how the watermarks will work.
The company’s commitment to transparency and compliance with regulations is a positive step forward in the development of AI-generated content.