
What happened
Anthropic is embedding imperceptible, machine-readable watermarks into text generated by Claude models released on or after August 2. The watermark travels with copied text and is designed to survive some editing, though heavy rewrites or translations may remove it. Images will also receive watermarks to show Claude processed them and flag tampering.
Why it matters
The move addresses the EU AI Act's transparency rules that took effect August 2, requiring generative AI providers to make synthetic output machine-readable and detectable. It also reflects a broader industry and user backlash against low-quality AI-generated content (called "AI slop") flooding social feeds, with platforms like YouTube and Substack rolling out their own detection tools.
What to watch
The watermark only indicates Claude had a hand in something, not that it generated the entire text—even proofreading or translating a paragraph leaves a trace. Anthropic notes the detection works for current Claude models, but the company is still working on extending it to older models. Watermarking alone won't prevent determined users from erasing text watermarks, which have historically been easy to remove.
Summaries like this, in your inbox every morning.
Anthropic's watermarking initiative sits at the intersection of regulatory compliance and platform pressure to curb AI-generated spam. The EU AI Act's transparency rules, which took effect August 2, created a legal obligation for the company to make synthetic output detectable—a requirement Anthropic is meeting by embedding signals at the model level rather than at the application layer. This ensures the watermark persists across all deployment channels: the chatbot, API, and third-party tools like Claude Code.
The timing also reflects a broader reckoning with what the industry calls "AI slop"—low-quality, often mass-produced AI-generated content that has flooded social media feeds and search results. YouTube has tightened its "inauthentic content" policy to deny monetization to channels leaning on generic, templated AI output; Substack has introduced reader-triggered AI scanners to flag the ratio of human to machine-written text. Anthropic's move positions the company as aligned with these efforts, even though watermarking has historically been fragile—previous text watermarks have been easy to remove or degrade.
The approach introduces a meaningful trade-off: the watermark signals only that Claude participated in generating text, not that it created the entire output. A journalist using Claude to translate a transcript or a writer proofreading a paragraph both leave a watermark trace, potentially conflating legitimate AI assistance with mass-produced disinformation. As commentators have noted, a flat "AI" label risks treating these use cases identically, which may prove a difficult balance as sentiment toward AI-generated content hardens.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
Microsoft launched MAI-Transcribe-2-Streaming, its first streaming transcription model, priced at 54 cents per…
Microsoft AI said it released MAI-Transcribe-2-Streaming, which returns provisional results in just over 100 m…

Google announced Gemini 4 Argon on September 30, saying DeepMind's own evaluation beat GPT-6 Astra, Claude Fab…

OpenAI dismissed three researchers, according to reports, after highly confidential information was shared wit…

Anthropic PBC reportedly aims to begin marketing its IPO the week of Nov
On October 1, OpenAI updated ChatGPT's release notes with shopping features — a 'try on' button on product car…
