AIToday
Large Language ModelsAI Regulation & PolicyAI Safety & AlignmentITmedia AI+Published: Aug 16, 2026, 10:00 JST3 min read

Anthropic explains invisible watermark in Claude text

Anthropic explains invisible watermark in Claude text

Key takeaway

  • Anthropic has explained how it embeds invisible watermarks in Claude-generated text to comply with EU AI Act transparency rules.

  • The watermark works by replacing the random process that selects words with one governed by a secret key and context, leaving detectable statistical patterns without adding visible text or affecting performance.

  • The company is applying the watermark globally, not just in the EU, and will offer an API for detection, though the watermark can be erased by complete rewriting and does not work well on formulas, strict code, or single-answer facts.

3 Key Points

  1. What happened

    Anthropic revealed how it embeds imperceptible electronic watermarks in text generated by Claude. The method, based on Google DeepMind's SynthID-Text, replaces the random number source used to select words with a secret key and preceding context, leaving a statistical pattern. The watermark adds no characters, so it does not affect text quality, speed, or cost.

  2. Why it matters

    Anthropic signed the EU AI Act's transparency code of conduct, requiring disclosure that AI-generated text contains watermarks. Although the regulation applies only in the EU, the company is rolling out the watermark globally across all models and products, including in Japan, because it lacks a reliable way to limit implementation by region. Other major AI developers are expected to implement their own watermarks under the same obligation.

  3. What to watch

    Detection requires an API Anthropic will provide soon, using a secret key to verify word sequences and measure Claude's involvement. The watermark works weakly in low-choice contexts (formulas, code, facts with one correct answer) and is defeated by complete rewriting, but survives light edits. Anthropic will deploy the feature to existing models over the coming months before the regulation's August 2, 2026 compliance deadline.

Ask the AI about this article →

Context & Analysis

Anthropic's decision to apply the watermark globally, rather than limit it to the EU, reflects a practical compliance challenge: the company lacks a reliable technical method to restrict the feature by region. Under the EU AI Act's transparency code of conduct, which Anthropic and roughly 190 other organizations signed, developers must disclose when AI-generated text contains watermarks. Rather than engineer separate systems for different markets, Anthropic chose to deploy the feature worldwide, including Japan, making the watermark a global standard for its Claude product.

The watermark's design rests on a fundamental constraint of language generation: when an AI model chooses the next word, most word choices do not meaningfully change the meaning or readability of the text. By embedding a secret pattern in those low-risk selections—without adding extra text or characters—Anthropic avoids the common criticism that watermarking slows AI or degrades output quality. The transparency comes at minimal cost. However, the watermark has clear limits: it cannot survive a complete rewrite, and it fails to embed itself in contexts where only one word makes sense, such as finishing a historical fact or completing a mathematical formula. For strict code, Anthropic limits watermarking to comments, where alternative wording is possible.

FAQ

Does the watermark slow down Claude or increase costs?
No. Anthropic states that the watermark adds no characters to text, requires no extra tokens, and has only minimal impact on speed, so the price does not increase.
Can people remove or detect the watermark themselves?
The watermark is imperceptible to humans and survives light edits, but Anthropic says a complete rewrite replacing all words will erase it. Detection requires an API that Anthropic will provide soon, and requires the secret key.
Does the watermark work on all types of content Claude generates?
No. The watermark works poorly on text with few word choices, such as factual statements with one correct answer, mathematical formulas, and strict code where word choice affects function. It works better on translations, which use all words Claude selects, and on comments within code.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Nvidia revives Rubin CPX chip with major redesignYahoo Finance AI · 2h ago
  • AI advice followed by 79%, but well-being unchangedITmedia AI+ · 5h ago
  • Enterprises face agent governance gapSiliconANGLE AI · 8h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleOpen-weight AI models threaten pricing, but cloud giants stay profitable