guide

Which Claude Models Are Watermarked? The Rule, and Why Nobody Can Verify It Yet

5 min read
By Dr. Sarah Chen
Trusted by 2.5 million+ users
99.8% Success Rate
Free & Unlimited
99.8%
Bypass Rate
2.5 million+
Users Served
50+
Languages
Free
Unlimited Use

The Rule, Stated Precisely

There is a lot of confused reporting on this, including sites that publish confident model-by-model lists which contradict each other. Here is the only claim traceable to Anthropic itself:

Models launched on or after 2 August 2026 support text watermarking at launch. Support for older models is "in progress".

That is a rule, not a list. Applying it correctly matters more than memorising which model was covered on any given day, because the "in progress" category can change without a separate announcement.

Model launch dateWatermark statusConfidence
On or after 2 Aug 2026Enabled at launchHigh — stated by Anthropic
Before 2 Aug 2026"In progress" — may be enabled during the transitionLow — can change silently
Any model, output generated before enablementNo markHigh — watermarks apply at generation

Why the Model Lists You See Are Unreliable

Three problems make third-party lists untrustworthy right now.

The transition period is undated. Anthropic says older-model support is "in progress" without committing to a completion date. A list accurate on 13 August may be wrong on 20 August.

No detector exists. Anthropic has not published its technical documentation or a detection tool. Nobody outside Anthropic can test a model's output and confirm whether a mark is present. Every confident public claim about a specific model is therefore inference, not measurement.

Coverage is per-product as well as per-model. Anthropic lists Claude, the API, Claude Code, Claude Cowork and Claude Tag as covered surfaces. A model reached through a cloud partner or an unsupported file type may behave differently.

We could publish a model table here and it would get traffic. It would also be a guess presented as fact, which is the opposite of useful.


What "Retroactive" Actually Means

This comes up constantly, and the answer is reassuring and technical.

A text watermark is applied during generation. The model biases its token selection as it writes, producing a statistical pattern. There is no post-processing step, and no way to reach into text that already exists and add one.

So:

ScenarioCarries a mark
Text generated before that model had watermarkingNo
Text generated after enablementYes
Old text pasted into Claude and returned unchangedYes — the returned copy is new output
Old text you edited yourself in your own editorNo

That third row is the one people miss. If you paste a 2025 document into Claude and it hands text back to you, what you copy out is freshly generated output from a covered model, regardless of when you originally wrote it.


What You Can and Cannot Verify Today

QuestionAnswerable now?
Does Anthropic watermark text?Yes — confirmed by Anthropic
Which products are covered?Yes — Claude, API, Claude Code, Cowork, Tag
Is a specific model covered?Only by launch date, not by testing
Does this specific paragraph carry a mark?No — no public detector exists
What does the mark look like?No — technical spec unpublished
Will a journal or university detect it?No — they have no tool either

That last row deserves emphasis, because it drives a lot of anxiety that is currently unfounded. Institutions cannot check for a Claude watermark today, for the same reason you cannot: Anthropic has not shipped the detector. That will change, and Anthropic has said documentation is forthcoming.


Where the Other Labs Stand

ProviderText watermarkingMechanism
Anthropic (Claude)Deployed from 2 Aug 2026Proprietary, spec unpublished
Google (Gemini)Deployed in supported productsSynthID
OpenAI (ChatGPT)Researched, not publicly deployed as of Aug 2026

The forcing function is Article 50 of the EU AI Act, in force since 2 August 2026, which requires providers of generative AI to mark output in a machine-readable form wherever technically feasible. Anthropic and Google have moved. Pressure on the rest is regulatory rather than voluntary, which is why this is unlikely to reverse.


What To Do With This

If you are trying to work out whether your text is marked: check when the model you used was launched. If it was before 2 August 2026, treat the answer as genuinely unknown rather than "no" — Anthropic's transition language allows enablement without notice.

If you are worried about being detected: nobody can detect it yet. When tooling ships, the same limits Anthropic already documents will apply — the mark does not survive heavy editing, paraphrasing, translation or mixing with other writing, and short passages carry too little signal.

If you are setting institutional policy: do not write a rule that depends on detecting a watermark. There is no tool to enforce it with, and when one arrives, Anthropic's own guidance is that a detected mark shows content "may have been processed by Claude" and is "not fully conclusive" about authorship.

For the full mechanism, see what the Claude watermark actually detects. If your own writing carries a mark because you used Claude as an editor, read Claude watermark false positives.

DSC

Dr. Sarah Chen

AI Content Specialist

Ph.D. in Computational Linguistics, Stanford University

10+ years in AI and NLP research

FAQ

Frequently Asked Questions

Anthropic states that models launched on or after 2 August 2026 support text watermarking at launch, and that support for older models is "in progress". So the answer for any given model depends on its launch date relative to 2 August 2026 — and for models released before that date it can be switched on during the transition period without a separate announcement.

For covered models, yes — Anthropic describes it as applied at the model level, worldwide, across Claude, the API, Claude Code, Claude Cowork and Claude Tag. What you cannot currently do is verify it on a specific piece of text, because Anthropic has not published detection tooling or the technical specification.

No. A watermark is applied during generation, so it cannot be added to text that already exists. Output produced before a model had watermarking enabled carries no mark, and nothing Anthropic does later changes that text.

Right now you cannot check the text itself, because no detector has been published. The only reliable method is the rule: compare the model's launch date to 2 August 2026, and treat any pre-August model as "may be enabled at any time" per Anthropic's transition language.

Anthropic lists Claude Code among the covered products. The practical picture for code is less clear than for prose: watermarks of this type rely on having enough text to carry a statistical signal, and Anthropic states short passages may not contain enough for reliable detection.

Google marks text in supported Gemini products using SynthID. OpenAI has researched text watermarking but had not publicly deployed it for ordinary ChatGPT output as of August 2026. The EU AI Act's Article 50 transparency obligations, in force since 2 August 2026, apply pressure across all providers.

Ready to Humanize Your Content?

Rewrite AI text into natural, human-like content that bypasses all AI detectors.

Instant Results
99.8% Bypass Rate
Unlimited Free