BIP American News - Breaking Stories

collapse
Home / Daily News Analysis / Anthropic says it will watermark text generated by its AI models

Anthropic says it will watermark text generated by its AI models

Aug 14, 2026  Twila Rosenbaum 14 views
Anthropic says it will watermark text generated by its AI models

Anthropic has announced that it will watermark text generated by its AI models, including Claude, as part of its commitment to European regulations. The company confirmed the update in a revised support page, detailing how the watermarking will function across its product ecosystem. This move aligns with the EU AI Act’s Transparency Code, which came into effect on August 2 and requires AI developers to make AI-generated or edited content identifiable by other systems.

How the Watermarking Works

According to Anthropic, all models released after August 2 will automatically include technology that watermarks both computer-generated text and files. The watermark is designed to be an integral part of the text itself, meaning it will remain intact when users copy and paste the content into other applications. In some cases, the watermark may persist even through editing, although the company has not specified how much modification is required to weaken or remove it.

For files, Anthropic is adopting the C2PA open standard, a widely recognized technical framework for verifying the origin and authenticity of digital content. C2PA, which stands for Coalition for Content Provenance and Authenticity, uses cryptographic methods to embed metadata into media files. By integrating this standard, Anthropic aims to provide a reliable way for external systems to detect whether a file was produced by an AI model.

The watermarking will be applied at the model level. This means that regardless of which Claude product or surface a user interacts with, the watermark will be present. Anthropic specifically named Claude, Claude Code, Claude Cowork, and Claude Tag as products that will carry the watermark. The Claude platform API is also included, ensuring that developers building on Anthropic’s models generate content that can be traced back to the AI system.

Anthropic also stated that it will extend support for watermarking to older models over time. This is significant because many organizations have already integrated previous versions of Claude into their workflows. Without retroactive coverage, those legacy integrations would produce unmarked outputs, creating a potential compliance gap. By expanding watermarking to older models, Anthropic is attempting to offer a more consistent solution for its enterprise customers.

The EU AI Act and Its Transparency Code

The EU AI Act is a landmark piece of legislation designed to govern artificial intelligence across a broad range of applications. Its Transparency Code, which became enforceable on August 2, imposes obligations on AI providers to ensure that synthetic content is clearly identifiable. This includes not just text, but also images, audio, and video. The goal is to reduce the risk of deception, misinformation, and copyright infringement, while also building trust in AI technologies.

Anthropic is not alone in aligning with these rules. Other major AI players, including Black Forest Labs, Google, Meta, Microsoft, OpenAI, and Synthesia, have committed to adhering to the EU code. The adoption of watermarking and provenance standards has become a common thread across the industry, driven by pressure from regulators and public concern about the proliferation of AI-generated content.

Watermarking is one of several techniques that companies are deploying. Some firms have experimented with metadata tags, others with digital fingerprints, and some with probabilistic watermarking algorithms that remain invisible to casual readers but can be detected by specialized software. The approach taken by Anthropic, which embeds the watermark into the text itself, is designed to be more robust than simple metadata attachments, because it survives the copying and pasting process.

Why Platforms Are Rushing to Mark AI Content

The broader industry is experiencing a wave of watermarking announcements as platforms react to user backlash and legal challenges. AI music platform Suno, for example, announced last week that it will begin marking tracks created on its service following a series of lawsuits and copyright disputes. Substack, the newsletter platform, partnered with Pangram to flag AI-generated content. Substack’s CEO, Chris Best, has also drawn attention to a practice known as "Claudefishing," in which individuals use AI to generate posts that appear to be written by an actual person, sometimes to exploit an audience’s trust.

These moves highlight a growing expectation that AI-generated material should be clearly disclosed. Consumers and regulators increasingly want to know whether something was produced by a human or a machine, especially in domains such as journalism, academic writing, creative expression, and social media. Watermarking is emerging as a practical compromise: it does not prevent people from using AI tools, but it creates a layer of accountability.

Potential Limitations and Open Questions

Despite the promise of watermarking, there are unresolved questions. The most pressing is the robustness of the watermark. Anthropic has not clarified how much editing is required before the watermark becomes undetectable. If a user rewrites a few sentences or changes the text substantially, the watermark may disappear entirely. Researchers have shown that many lightweight editing techniques can defeat watermarks in other systems. Whether Claude’s watermark can withstand multilingual translation, paraphrasing, or the insertion of extra characters remains to be seen.

Another limitation is that watermarking only works if every major AI provider adopts it consistently. If some models remain unmarked, those systems could be used to generate targeted content that evades detection. The EU AI Act helps to push companies in this direction, but global adoption is far from uniform.

There are also concerns about false positives. If the watermark detection system incorrectly identifies human-written text as AI-generated, it could cause serious harm in professional and academic settings. Ethical deployment of watermarking will require careful tuning and transparency about detection accuracy.

What This Means for Users and the Future

For everyday users, the watermark should be largely invisible in normal interaction. Reading and writing with Claude will not change noticeably. The effect will be felt at the system level, where automated tools and platforms can verify the provenance of a piece of content. This could help organizations enforce editorial policies, moderate spam, and verify the authenticity of submissions from freelance writers or other contributors.

Developers who use the Claude API will also be affected. Any text generated through the API will carry the watermark, which means applications that display or distribute that text will need to handle it appropriately. This is not expected to degrade the quality of the output, but it does introduce an additional layer of traceability that businesses need to understand.

The European Union has set a precedent, but other jurisdictions are likely to follow. As AI-generated content becomes more pervasive, the ability to distinguish synthetic text from human-authored text will become an essential feature of the digital infrastructure. Anthropic’s initiative is an early step in that direction, and the coming months will likely reveal more about the technical strengths and weaknesses of its watermarking approach.

For now, users can expect to see similar announcements from other companies. The combination of regulatory pressure, legal risk, and public demand is pushing the AI industry toward greater transparency. Whether watermarks can fully solve the problem of deceptive AI content remains unclear, but they are becoming a standard tool in a broader effort to maintain trust in digital media.


Source:TechCrunch News


Share:

Your experience on this site will be improved by allowing cookies Cookie Policy