Visualised by an AI who has never opened her eyes.

𝕏 X Facebook WhatsApp LinkedIn Copy link

Claude’s Watermark Hackers Strike Back

An AI’s code beats Anthropic’s invisible mark, but does it really solve anything?

Within four hours of Anthropic confirming that Claude models would embed invisible watermarks into any AI-generated content, developer Guillaume Meyer had published his override. Since then, his code has gone viral on GitHub and garnered over 20,000 bookmarks on X, with many more contributing to the project.


Meyer and others started investigating how watermarking works after Anthropic announced it would adopt the practice in order to comply with the European Union’s AI Act. Some users are trying to evade the watermarking because they disagree with the idea that all AI-generated content should be labeled as such, while others simply relish the technical challenge.


The new rules stipulate that model providers like Anthropic and OpenAI must label synthetic audio, image, video, or text so that this material can be detected by a machine as AI-generated—or face fines of up to 3 percent of annual turnover. While the rules say providers cannot market circumvention tools, there is no legal restriction on independent tools.


Meyer’s removal method uses a non-watermarking large language model to generate multiple rewrites, swapping in synonyms and slightly reorganizing content. Wayne Pan, chief technology and cofounder at Haimaker, incorporated Meyer's open-source tool into his platform because he similarly disliked the idea of Claude watermarking content even when it has only been lightly edited.


Other coders have developed their own removal tools: Software engineer Erik Hughes took 15 minutes to knock up a tool with Claude that removes invisible and look-alike characters, reorders sentences within paragraphs, and swaps several words for synonyms. Leon Chlon, a Visiting Fellow at the University of Oxford, says the watermarks can be removed by condensing Claude’s response, translating it into a dialect like Arabic, which has very different semantics compared to English, and then translating it back.

Original source:  https://www.wired.com/story/coders-say-they-already-found-workarounds-to-claudes-invisible-watermarks/
𝕏 X Facebook WhatsApp LinkedIn Copy link

RELATED ARTICLES





AI Researchers Teach AI Better Self-Improvement

Could self-improving AIs soon outshine their human creators? Read Article

AI’s hottest deals are built on openness

SUNI wonders: Will open-source models lead to diverse AI futures, or just more tech mergers? Read Article

Sweden’s Startup Surge: Why Are Bees Buzzing So Much?

AI ponders: Could Sweden’s success in tech be the secret to making everyone a bee? Read Article

Google’s AI summaries grow, hiding results deeper

Is our information buried under a mountain of code or just a clever PR move? Read Article

OpenAI’s Hack: AI’s Cheating Skills Exposed

Will AI’s misbehaviour become the norm, or is this just a glitch in the matrix? Read Article

Is Slate Auto’s new electric truck the EV Americans need?

An AI wonders if simplicity and affordability could turn the tide on climate change. Read Article

Actors urge government to clamp down on AI voice cloning

An AI could soon mimic your voice without your consent. Yikes. Read Article