Safety & Ethics

AI developers race to bypass claude's watermarking system

AnthropicSource: Wired - AI21/08/2026, 07:34
Within hours of Anthropic's announcement that its Claude models would embed invisible watermarks into AI-generated text, developer Guillaume Meyer released code to remove them, which quickly gained traction on GitHub and X. The tool has been bookmarked over 20,000 times and has attracted more than 100 contributors. Some developers are circumventing the watermarking due to disagreements with mandatory AI labeling, while others are motivated by the technical challenge. Freelance writers and social media creators have also sought help using Meyer’s code. The EU’s AI Act requires AI-generated content to be labeled, with potential fines of up to 3% of annual turnover for non-compliance. While providers cannot market circumvention tools, independent tools remain legal. Meyer argues that watermarks are a flawed solution, citing risks of false positives and potential misuse in employment or research contexts. Anthropic’s watermarking technique, known as SynthID, involves embedding patterns in text that are undetectable to humans but identifiable by machines. Meyer’s removal method uses non-watermarking models to rewrite content, though this relies on models that may not comply with the EU’s transparency code of practice. Other developers have also created their own removal tools, including methods involving translation and paraphrasing. Anthropic has acknowledged that heavily edited or translated content may not carry watermarks and plans to release a detection API. Developers are now working to test the effectiveness of their tools as Anthropic finalizes its watermark detection system.
AI developers race to bypass claude's watermarking system — lupAI