AIToday
Large Language ModelsAI Regulation & PolicyImpress WatchPublished: Oct 6, 2026, 16:00 JST

OpenAI to watermark ChatGPT and Codex text with textGrain in EU

OpenAI to watermark ChatGPT and Codex text with textGrain in EU

3 Key Points

  1. What happened

    OpenAI is adding the textGrain watermark to text from ChatGPT and Codex, applying it to EU outputs over the coming weeks. From October 5, API users worldwide can switch it on for some models, off by default.

  2. Why it matters

    The move responds to the EU AI Act's requirement that providers make AI-generated text machine-detectable, so outputs may carry a signal firms and regulators can check — though OpenAI says detection has limits.

  3. What to watch

    OpenAI will initially give the detector only to approved researchers and institutions, not the public, so real-world usefulness hinges on who gets access. Its own tests found detection drops to 66% when 10% of words are swapped.

WHO IT HITSCompliance, trust-and-safety, and platform teams at companies shipping generative AI products into the EU will need to account for a watermark that ships by default on EU outputs and is optional in the API. Editing tools and detection-reliant workflows may be affected, since OpenAI's tests show light rewording hurts detection.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The rollout is tied directly to the EU AI Act, which requires generative AI providers to make AI-generated text machine-identifiable. OpenAI is answering that with textGrain, a technique that avoids stamping characters or metadata onto the text and instead bakes a statistical signal into how the model picks its words.

OpenAI's own numbers show where the approach holds and where it frays. Detection is strongest on prose with wide word choice, but weakens when the text is short, when the subject offers few alternatives, or when a person edits it — swapping 10% of words in a 400-token English passage cut detection from about 92% to 66%, and swapping 25% dropped it to 17%. The company also stresses that a detected watermark cannot establish human involvement, ownership, responsibility, or user identity, and that a missing watermark does not prove human authorship.

That is why the detector will go first to approved researchers and institutions rather than the general public. The practical stakes look likely to hinge on how much access those outsiders get, and whether detection holds up outside OpenAI's tests — especially on text that has passed through editing tools.

FAQ
How does textGrain watermark text?
It does not add special characters or metadata. Instead, it injects an invisible statistical signal into the word choices the model makes while generating, which a dedicated detector can read.
How accurate is the watermark?
OpenAI's evaluation found that at a 1% false-positive rate, detection hit about 80% at 200 tokens and about 95% at 400 tokens for prose like psychology. For math, where word choice is limited, detection dropped sharply.
Can a detected watermark prove who wrote the text?
No. Even when detected, it cannot show how much a human was involved, who owns the text, who is responsible, or the user's identity. It also does not guarantee accuracy.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleDeepMind: Chinchilla at 70 billion parameters beats Gopher at 280 billion