Back to feed
Artificial intelligenceSingle source

OpenAI details invisible text watermarking for ChatGPT, Codex and API

OpenAI announced textGrain, a system for detecting text generated by its models. Eligible ChatGPT and Codex outputs in the EU will receive watermarks in the coming weeks, while API customers globally can opt in for selected models starting today.

Preferred on Google
Text size

Image: 9to5mac.com · Author: Marcus Mendes · source articleEditorial excerpt for news reporting

OpenAI announced textGrain, an invisible text-watermarking system intended to help identify material generated by its models. The company presented the system in the context of the EU AI Act’s requirements for identifying AI-generated content.

API customers globally can opt in to watermarking for selected models starting today. OpenAI plans to add it to eligible ChatGPT and Codex text outputs in the European Union in the coming weeks.

How textGrain works

textGrain adds an invisible statistical signal to a model’s word choices. A detector searches for that signal and assesses whether a passage contains an OpenAI watermark.

OpenAI has opened applications for detector access. Initially, only approved researchers and expert organizations will be able to use it while the company evaluates and improves the technology. OpenAI also plans to release the system as open source.

Detection limits

OpenAI warns that the detector can produce false positives and false negatives, particularly for short passages and domains with less flexibility in word choice, such as mathematics.

In a test of 400-token passages, replacing 10% of words with synonyms reduced detection from about 92% to 66%; replacing 25% reduced it to 17%. Editing and translation can also weaken the watermark.

The watermark does not measure human contribution, establish ownership or responsibility, identify the user, or verify accuracy. OpenAI also says that an undetected watermark does not prove human authorship.

The company said it will revisit the approach as the technology, standards, and evidence develop, including research into whether watermarking can better distinguish AI assistance from AI authorship.

What we know

  • OpenAI introduced textGrain to statistically watermark text generated by its models.
  • API customers can opt in for selected models starting today.
  • ChatGPT and Codex watermarking will initially cover eligible text outputs in the EU.
  • Replacing 25% of words with synonyms reduced detection from 92% to 17% in a test.
  • The watermark does not establish human authorship, ownership, or the user’s identity.

Trust: Single source

  • The story relies on one publisher. A second independent confirmation has not been established from the cited sources.
  • Cited links: 1. Publisher groups: 1. Sources: 9to5Mac.
  • Monitoring in the past hour: 37 of 37 RSS sources; available: 36, unavailable: 1. Matching covers published stories from the past 7 days. This does not cover the entire web.
  • Additional citations require matching headlines, context, event timing and accessible article text. This automatically finds related coverage; it does not establish independence or the truth of every claim.
Related updates appear in the story timeline. This assessment changes with the sources attached to this story.
View sources1

COMMUNITY

Discussion

0

No comments yet. Start the discussion.