OpenAI 为遵守欧盟 AI 法案,将在未来数周对欧盟符合条件的 ChatGPT 和 Codex 文本加入隐形水印,API 客户现已可对部分模型开启。
OpenAI will invisibly watermark ChatGPT and Codex text in the EU.
The watermark goes only into text that ChatGPT or Codex writes. So if someone in the EU asks ChatGPT for a paragraph and pastes it into a personal blog, the hidden signal travels with those words, because it lives in the pattern of word choices rather than in any special characters.
If that person rewrites a lot of it, the signal weakens and may disappear. Anything a person writes on their own carries no watermark at all.
Eligible users on every plan get the watermark over the coming weeks, while API customers worldwide can switch it on today for select models.
the public can't use OpenAI's tool for checking whether a piece of text carries the watermark
The rollout answers Article 50 of the EU AI Act, and generative systems already on the market before August 2 must carry machine-readable marks by
Nobody can see it by looking, because the text reads exactly like normal text: no marks, no odd characters, nothing a reader, editor or teacher would notice.
The only way to find it is to run the text through OpenAI's detector, and right now OpenAI keeps that tool to itself plus a small group of researchers and expert organizations who apply and get approved one by one.
The scheme, textGrain, tilts the model's word choices to leave a statistical signal invisible to readers but measurable by a detector.
At a 1% false positive rate, that detector caught about 80% of 200-token psychology passages and about 95% of 400-token ones in OpenAI's tests.
Mathematics answers scored substantially lower, because tight wording leaves fewer interchangeable words to carry the signal.
Editing erodes it too: swapping 10% of words for synonyms in 400-token passages cut detection from about 92% to 66%, and swapping 25% cut it to 17%.
A positive result identifies no user, prompt or owner, while a negative one proves little, since short, edited, translated or rival-generated text escapes detection.
We're expanding our approach to content provenance to include text in response to EU regulatory requirements, while recognizing the significant limitations of current text watermarking technology. Our tools already help verify whether an image or audio file was created with our models. This work builds on those efforts to help people better understand when content may have been generated or edited with an OpenAI model. In the EU, we’ll start watermarking eligible text from ChatGPT and Codex over the coming weeks to comply with the EU AI Act. Customers using our API can turn on text watermarking for select models worldwide today.在 X 查看被引用的帖子
来源:Rohan Paul · x.com