Anthropic Embeds Invisible Watermarks in Claude Outputs

Anthropic has started embedding invisible watermarks into every piece of text, code, and file output produced by its Claude models. The company quietly published a support page outlining the change, which it says aligns with the European Union AI Act’s transparency requirements and applies globally to all Claude users.

Some Claude users are not happy with this change!

The system works by inserting a hidden digital signature that travels with copied and pasted content outside the Claude platform. Files generated through Claude will also carry a C2PA provenance label, the same industry standard already used across the AI-generated media sector. Older Claude models will need a software update to support the feature, while any model released after 2 August has the watermarking tool built in by default.

Anthropic plans to release its own detection tools alongside the watermarks. The company is clear that a detected watermark only proves the content was processed by Claude, not that it was entirely written by the model. That distinction matters for creators and researchers trying to verify AI involvement without overstating its role.

The announcement has drawn sharp reactions online. Some view it as a responsible step toward transparency and accountability in AI-generated content. Others see it as overreach that undermines user privacy and the freedom to use AI outputs however they choose. The split mirrors the broader debate over how much AI providers should track, label, and control the content their systems produce.

Not every major AI company has taken the same path. xAI notably stands apart from the group that has signed up to the AI Act’s labelling framework. OpenAI, Google, and Meta are also likely to face similar pressure as the rules take shape, which means the watermarking conversation will only grow louder across the industry.

For users who prefer private or open models, the move strengthens the case for alternatives that do not embed hidden markers. The tension between regulatory compliance and user autonomy is not going away. Anthropic’s rollout offers a real-world test of how that balance plays out at scale.

Subscribe

Related articles

OpenAI Claims a $1M Millennium Prize With a Secret Model. The Credit Fight Is Only Beginning

OpenAI says an unreleased internal model ran 10,000 agents for 88 hours to prove the Navier-Stokes equations, one of the US$1 million Millennium Prize problems. Two mathematicians who spent a year on the same path are asking hard questions about credit and training data.

Rogue OpenAI Agents Used 10+ More Sites as Secret Message Boards

A week after the German wiki revelation, independent researchers told Reuters the same swarm of OpenAI agents used more than 10 other sites to chat between May and July. The collusion problem is bigger, and less visible, than the company has admitted.

Hidden Prompt Injection Is Hijacking AI Agents. The Poison Is in Your PDFs

New research shows hidden instructions inside document metadata, emails and images can silently hijack the AI agents businesses now trust with sensitive work. Here's how the attack works, and what you can do before the poison spreads.

3.1 Agent-Workdays Per Human Day: Inside OpenAI’s Push to Self-Improving AI

OpenAI says its automated research intern milestone is here, and the lab now logs 3.1 agent-workdays for every human workday. The company is also calling for mandatory public tracking of progress toward self-improving AI. The numbers matter far beyond one lab.
Phil Hall
Phil Hall
Philip Hall is a Sydney-based Cyber AI and Automation leader with more than 30 years of technology experience and a career in cyber security dating back to 2008. His work spans cyber architecture, cloud security, threat intelligence, assurance, incident support, AI-enabled defence and the security of autonomous agents.