ChengRang

OpenAI textGrain

AI Content Detection Free

A statistical text watermarking scheme announced by OpenAI on October 5, 2026. It adds an invisible statistical signal through the model word choices so a detector can tell whether text was generated or processed by an OpenAI system. API customers worldwide can opt in for select models from day one, EU ChatGPT and Codex text gets the watermark over the coming weeks, and detector access starts with approved researchers and expert organizations

Text WatermarkingContent ProvenanceEU AI ActComplianceOpenAIContent Labeling
Visit OpenAI textGrain

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

textGrain is a statistical text watermarking scheme OpenAI announced on October 5, 2026, laid out in a post titled Our approach to EU text provenance rules. It adds an invisible statistical signal through the model word choices: at every step of generation there is a cluster of near-equally likely candidate tokens, and a secret key combined with the preceding words quietly decides which of those candidates gets favored. No single word looks odd, but across a long enough passage the pattern becomes measurable. A detector holding the same key re-derives the preferred options at each position, counts how often the text followed them, and compares that rate to what chance would produce. A large gap means the text very likely came from the watermarked model.

Regulation drove this, not product strategy. Article 50 of the EU AI Act requires systems that generate synthetic text to mark their output in a machine-readable way, and reporting around the announcement puts the start of those transparency rules at August 2, 2026. That is why the rollout begins in the EU and why OpenAI framed the post as its approach to EU text provenance rules.

Key Features

Use Cases

Pros

Pricing

textGrain carries no separate charge. API usage keeps its existing token-based billing, and turning the watermark on does not change pricing. ChatGPT and Codex output includes it for EU users at no extra cost on their existing plans. The detector is not distributed publicly; approved researchers and expert organizations apply case by case, and there is no paid tier for general users.

Summary

textGrain reads more accurately as a compliance instrument than as a detector. It answers the Article 50 requirement for machine-readable marking and offers a usable provenance signal along the way.

Its strengths and weaknesses both follow from the mechanism. Longer text carries more signal: OpenAI's published figures put detection at roughly 80% around 200 tokens and 95% at 400 tokens. Editing erodes it fast, with about 10% of words swapped dropping detection from 92% to 66%, and roughly a quarter swapped leaving about 17%. Translation replaces the word choices entirely and likely destroys the signal, while code and structured output have too little entropy to carry much in the first place. Those numbers come from OpenAI and have been relayed through secondary coverage; no third party with detector access has published an independent reproduction yet.

One point is easy to miss. The watermark indicates whether text was generated or processed by an OpenAI system, not whether it was generated by AI. Text a model rewrote or polished also carries the signal, so a positive result does not mean the whole passage was machine-written, and a negative result is even weaker evidence of human authorship. OpenAI says both things explicitly.

If you serve AI-generated text in the EU, or you are weighing whether to build watermark checks into your own workflow, this is worth understanding now. Just do not ask it to carry conclusions it was not built to carry.

Version History

Category
AI Content Detection
Pricing
Free
Tags
Text Watermarking · Content Provenance · EU AI Act
Website

Related Tools