⚡
Brief IA
›

Anthropic Details Claude's Watermark and Code

💻 Code & Dev·Tom Levy·

Anthropic Details Claude's Watermark and Code

Anthropic Details Claude's Watermark and Code
⚡
Key Takeaways
1Anthropic plans to use SynthID-Text and a detection API to watermark Claude's text.
2The watermark withstands slight modifications, but a complete rewrite removes it, according to the company.
3The code will contain little watermarking, with a negligible effect on the produced code, according to Anthropic.
💡Why it matters — This clarifies how Anthropic intends to comply with the European transparency code and how detection will apply differently to text and code.
⚡Le brief IA que lisent les pros

Le brief IA que les pros lisent chaque soir

Les 7 actus IA du jour, décryptées en 5 min. Gratuit.

Inclus dès l'inscription : notre sélection des meilleurs guides & comparatifs IA.

Choisis ton rythme

Gratuit · Pas de spam · Désabonnement en 1 clic

📄
Full Analysis

Anthropic has clarified how the text generated by Claude will be watermarked, and under what conditions this marking remains detectable. The company plans to use SynthID‑Text and open a detection API, while ensuring that the output quality does not change. The framework invoked is that of the Transparency Code of the European AI Act, while user feedback continues to multiply.

Little watermarking in code and deployment by other actors

Anthropic indicates that the generated code will contain less watermarking than other types of text, as the model must produce functional code and has fewer possibilities for variation. However, the watermark may appear in parts of the code where arbitrary choices are possible, such as in comments. According to the company, this marking will have a negligible effect on the code actually produced. Anthropic also specifies that other major developers who have signed the same code of practice will implement their own watermarks.

Minor edits, complete rewrites: what the marking endures

Anthropic acknowledges that a text can be rewritten to mask the watermark: slight modifications will likely not completely remove it, while a complete rewrite where every word is replaced will erase it. In this case, the company believes it becomes debatable to qualify the text as AI-generated. For texts that are simply proofread or modified by Claude, the detectability of the watermark depends on the length and extent of the changes. If the edits are minimal, almost all the words remain those of the human author, and there is little, if anything, for the watermark to attach to.

Encoded pattern in low-stakes decisions, not a style detection

Anthropic explains that during low-stakes decisions, the model can insert an undetectable pattern for the reader but detectable by someone with the encoding key. The company distinguishes this process from AI detection methods that look for stylistic clues, such as those proposed by Pangram. According to Anthropic, spotting these writing patterns is fundamentally different from verifying a watermark.

SynthID‑Text and upcoming API, European compliance and reactions

Anthropic plans to use SynthID‑Text, an approach described in 2024 by the Google DeepMind team, and to publish a detection API. The company asserts that the quality of responses is not affected and that a watermarked text is indistinguishable from a non-watermarked text for a reader. The implementation is part of the Transparency Code of the European AI Act, which Anthropic has decided to comply with by announcing this project earlier in the week. The blog post, published on Friday, details the operation, resistance to modifications, and impact on the code. Reactions are mixed: on Reddit, some denounce a conspiracy, while others believe that refusing the watermark aims to deceive; it has been reported that dozens of users on X have canceled their subscriptions to Claude.

⚡

Brief IA — L'actualité IA en français

L'essentiel de l'actualité de l'intelligence artificielle, décrypté et expliqué chaque jour.