2 Sources
[1]
Anthropic pledges to embed watermarks to help discern AI slop in sop to EU
Anthropic will embed watermarks in the text and files generated by future models it launches in the EU, as part of its effort to comply with content and transparency rules in the bloc's AI Act. "Generated text will carry embedded watermarks, and generated files will include digitally signed provenance metadata where supported," the company said in a help document published on Monday. The AI biz is also working on models it has already released, to add output marking during the transition period allowed under EU law. "Marking will apply to output from supported models wherever Claude is offered, worldwide," the company said. The move may further amplify the appeal of open weight models and alienate Claude customers, who don't necessarily want consumers of their AI-generated content to know its provenance. In the past, Claude users have expressed frustration with pricing changes, reliability issues, and model safeguards that have hindered legitimate work in the name of safety. Claude users appear to be skeptical that a text-based watermarking scheme will work. Researchers have already demonstrated that image-based watermarking can be undone. Anthropic intends to apply marks to output from covered Claude models on the Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag. This also includes third-party providers of Anthropic models like AWS, Google Cloud, and Microsoft Foundry. The Ai biz expects to provide details about how people can detect Claude's marks, as required under EU law, in forthcoming documentation. "When a supported Claude model generates text, it weaves an imperceptible watermark directly into the text itself," the biz said. "You won't see it, and it doesn't change the meaning, quality, or readability of Claude's response. "Because the watermark is part of the text, it will travel with the text when it's copied and pasted elsewhere, and may persist through some editing. Watermarking will be applied at the model level, which means it will be present no matter which Claude product or surface the text comes from." Absent examples or technical documentation, it's unclear how Anthropic will make the watermarks hard to remove. But it should not be difficult to create an optical character recognition system that strips or omits obscure marks from Claude-generated text. Anthropic's insistence that its text marking scheme "doesn't change the meaning" of Claude's output appears to preclude using word choice and placement as a text provenance identifier - a technique Apple has reportedly used to catch those who would violate its employee secrecy agreements. As for marks attached to Claude-created files, Anthropic says it will rely on signed provenance metadata that conforms to the C2PA standard, for which there are already open source removal tools. Anthropic itself is already hedging about the utility of its marking method, noting that detected marks are not conclusive evidence that Claude produced the content and that the absence of marks cannot guarantee that AI wasn't involved in the creation of a particular piece of content. But perhaps the scheme is good enough to count as legal compliance. ®
[2]
AI watermarking: <b>Explained: Anthropic's plan to watermark all Claude-generated content, and how it works</b>
Invisible watermarks will be woven directly into text generated by Claude models. It will not alter the meaning, quality, or readability of a response and can travel with the text when copied elsewhere, potentially persisting through some editing. In a bid to make AI-generated material easier to identify, Anthropic has said it is working to add machine-readable marks to Claude-generated content. The move comes after the European Union's AI Act took effect. Anthropic has signed the Act's Article 50(2) Code of Practice on Transparency of AI-Generated Content, a voluntary framework aimed at making AI-generated material easier to identify. How it will work According to a post on Anthropic's website, Claude models launched in the EU on or after August 2, 2026 will support machine-readable marking at launch, with generated text carrying embedded watermarks and generated files having digitally signed provenance metadata where supported. Invisible watermarks will be woven directly into text generated by Claude models. It will not alter the meaning, quality, or readability of a response and can travel with the text when copied elsewhere, potentially persisting through some editing. For files generated via Claude, supported file types such as .svg, .png, or .jpg, will include signed provenance metadata. If a signed metadata label is present, it signals that a file was processed by Claude and lets you detect whether the file has been tampered with. Applicable to full Claude suite Anthropic said Claude markings will cover output from supported models, including Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag. Embedded watermarks will apply to all generated text. Provenance metadata will apply where Claude supports processing files. Conditions apply The AI giant cautioned that a detected mark does not mean it is a proof of AI authorship. "Claude may not be the original author. People often use Claude to proofread, translate, summarize, or convert files. The output can carry a Claude mark even if the underlying ideas, text, or data originated from another source," it said. Similarly, the absence of a mark also does not confirm that the content wasn't AI-generated, since older models, heavy editing, very short passages, or metadata stripped through format conversion could all prevent detection. EU's AI transparency rule in focus The European Union's new artificial intelligence transparency rules require companies to clearly label AI-generated content, including deepfakes, to help users distinguish between authentic and synthetic material. Under the new rules, companies operating AI systems, such as chatbots, must inform users that they are interacting with artificial intelligence. Text, images and other content generated using AI must also carry clear labels, which may include watermarks or other markers to facilitate identification. Existing AI systems have until December 2 to adapt to the new rules, and there are exemptions for "artistic, creative, satirical, fictional" work.
Share
Copy Link
Anthropic announced it will embed invisible watermarks in text and digitally signed provenance metadata in files generated by Claude models to comply with the EU AI Act. The watermarking will apply globally across all Claude platforms including API, AWS, Google Cloud, and Microsoft Foundry starting August 2026.
Anthropic has pledged to embed watermarks in all content generated by its Claude models as part of its effort to comply with transparency requirements under the EU AI Act
1
. The AI company announced that Claude models launched in the EU on or after August 2, 2026 will include machine-readable marking at launch, with generated text carrying embedded watermarks and generated files featuring digitally signed provenance metadata where supported2
. Anthropic has signed the EU AI Act's Article 50(2) Code of Practice on Transparency of AI-Generated Content, a voluntary framework designed to make distinguishing synthetic from authentic content easier for users.
Source: The Register
The watermarking approach involves weaving invisible watermarks directly into text generated by Claude models. According to Anthropic, these imperceptible marks will not alter the meaning, quality, or readability of Claude's responses
1
. The embedded watermarks will travel with the text when copied and pasted elsewhere, potentially persisting through some editing. For files generated via Claude, supported file types such as .svg, .png, or .jpg will include signed provenance metadata conforming to the C2PA standard1
. If a signed metadata label is present, it signals that a file was processed by Claude and enables detection of whether the file has been tampered with2
.The labeling of AI-generated content will apply to output from covered Claude models worldwide, not just in the EU. Anthropic stated that marking will be implemented at the model level across the Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag
1
. This includes third-party providers of Anthropic models such as AWS, Google Cloud, and Microsoft Foundry. The company is also working to add output marking to models it has already released during the transition period allowed under EU law. Anthropic plans to provide documentation on how people can detect Claude's marks, as required for legal compliance with the EU AI Act.Related Stories
Anthropic itself has acknowledged significant limitations with its marking system. The company cautioned that a detected mark does not serve as conclusive proof of AI authorship, as Claude may not be the original author when users employ it to proofread, translate, summarize, or convert files
2
. Similarly, the absence of a mark cannot guarantee that content wasn't AI-generated, since older models, heavy editing, very short passages, or metadata stripped through format conversion could all prevent detection2
. Claude users have expressed skepticism about whether text-based watermarking will work effectively, particularly since researchers have already demonstrated that image-based watermarking can be undone1
.The technical implementation raises questions about effectiveness in discerning AI-generated content from human work. Without detailed technical documentation, it remains unclear how Anthropic will make the watermarks difficult to remove. Experts suggest it should not be challenging to create an optical character recognition system that strips or omits obscure marks from Claude-generated text
1
. Additionally, open source removal tools already exist for the C2PA standard that Anthropic plans to use for file metadata. The move may amplify the appeal of open weight models and potentially alienate Claude customers who don't want consumers to know the provenance of their AI-generated content1
. Past frustrations among Claude users with pricing changes, reliability issues, and model safeguards suggest this new requirement could add to existing concerns about the platform's usability for legitimate work.Summarized by
Navi
24 Oct 2024•Technology

20 Nov 2024•Technology

19 May 2026•Technology

1
Technology

2
Science and Research

3
Technology
