AINEWS 搜索
返回 Rohan Paul (@rohanpaul_ai)
Rohan Paul (@rohanpaul_ai)· · 原发布时间

OpenAI 将为欧盟 ChatGPT 和 Codex 文本添加隐形水印

自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。

AI 辅助摘要

OpenAI 表示,未来几周将在欧盟为 ChatGPT 和 Codex 生成的符合条件的文本加入水印;全球 API 客户现可为部分模型开启。水印藏在遣词规律中,改写会削弱信号;公众目前无法使用 OpenAI 的检测工具。

正文 · 原文

OpenAI will invisibly watermark ChatGPT and Codex text in the EU.

The watermark goes only into text that ChatGPT or Codex writes. So if someone in the EU asks ChatGPT for a paragraph and pastes it into a personal blog, the hidden signal travels with those words, because it lives in the pattern of word choices rather than in any special characters.

If that person rewrites a lot of it, the signal weakens and may disappear. Anything a person writes on their own carries no watermark at all.

Eligible users on every plan get the watermark over the coming weeks, while API customers worldwide can switch it on today for select models.

the public can't use OpenAI's tool for checking whether a piece of text carries the watermark

The rollout answers Article 50 of the EU AI Act, and generative systems already on the market before August 2 must carry machine-readable marks by

Nobody can see it by looking, because the text reads exactly like normal text: no marks, no odd characters, nothing a reader, editor or teacher would notice.

The only way to find it is to run the text through OpenAI's detector, and right now OpenAI keeps that tool to itself plus a small group of researchers and expert organizations who apply and get approved one by one.

The scheme, textGrain, tilts the model's word choices to leave a statistical signal invisible to readers but measurable by a detector.

At a 1% false positive rate, that detector caught about 80% of 200-token psychology passages and about 95% of 400-token ones in OpenAI's tests.

Mathematics answers scored substantially lower, because tight wording leaves fewer interchangeable words to carry the signal.

Editing erodes it too: swapping 10% of words for synonyms in 400-token passages cut detection from about 92% to 66%, and swapping 25% cut it to 17%.

A positive result identifies no user, prompt or owner, while a negative one proves little, since short, edited, translated or rival-generated text escapes detection.

引用或回复的背景(作者 ID 4398626122,https://x.com/i/status/2107164650249101695): We're expanding our approach to content provenance to include text in response to EU regulatory requirements, while recognizing the significant limitations of current text watermarking technology.

Our tools already help verify whether an image or audio file was created with our models. This work builds on those efforts to help people better understand when content may have been generated or edited with an OpenAI model.

In the EU, we’ll start watermarking eligible text from ChatGPT and Codex over the coming weeks to comply with the EU AI Act.

Customers using our API can turn on text watermarking for select models worldwide today.

发现内容有误?提交纠错