OpenAI rolls out weak sauce watermarking for AI text
Compliance, we've heard of it
OpenAI has introduced a watermarking feature for its API customers' text output to comply with the European AI Act. The company plans to implement this watermarking, known as textGrain, in ChatGPT and Codex text generated within the EU. Unlike its competitor Anthropic, OpenAI will not apply the watermark to content produced outside of Europe, but API customers can still opt for the watermarking on all their generated content.
However, there are concerns about the effectiveness of this watermarking technique. The textGrain algorithm adds an invisible statistical signal to the model's word choices, making it possible for a detector to identify whether a passage contains an OpenAI watermark. The target error rate is one percent, but in practice, the detector may only flag 80 percent of the watermarks in short passages of 200 words, and even worse, in functional texts like math or code.
The watermark can be easily evaded by substituting words with synonyms, reducing detection rates significantly.
Written by urgent.news from The Register Science's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
This story
This is one outlet's version. Read the fullest account.
- OpenAI brings text watermarking to its API — and unlike Anthropic, it’s off by default thenewstack.io
- OpenAI details new text watermarking system for ChatGPT, Codex, and the API 9to5mac.com
- OpenAI will start watermarking ChatGPT’s text in the EU techcrunch.com
- OpenAI rolls out weak sauce watermarking for AI text theregister.com
- OpenAI is adding text watermarking in ChatGPT and Codex theverge.com
- To comply with the EU AI Act, OpenAI plans to add text watermarking for ChatGPT and Codex users in the EU and an opt-in setting for API customers globally (OpenAI) openai.com
- OpenAI is testing visual ads inside ChatGPT's image generation tool qz.com