AINEWS 搜索
返回 TechCrunch AI 报道
TechCrunch AI 报道· · 原发布时间 精选

OpenAI 将在欧盟为 ChatGPT 和 Codex 生成的文本添加隐形水印

自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。

AI 辅助摘要

OpenAI 称,未来几周将在欧盟向所有套餐的符合条件用户逐步推出文本水印;全球 API 开发者今天起可为部分模型手动开启,默认仍关闭。水印可随复制的文本保留,但改写会削弱检测:在一项测试中,替换 10% 的词后,检出率从约 92% 降至 66%。检测工具初期仅向获批研究人员和专业机构开放。

推荐理由

欧盟用户将遇到这一文本标记,但改写可显著降低检出率,说明水印检测结果有明确局限。

正文 · 原文

OpenAI will start adding an invisible watermark to text generated by ChatGPT and Codex in the European Union to comply with the EU AI Act, the company said Monday in a blog post. The EU AI Act’s transparency rules, which took effect on August 2, require AI companies to mark AI-generated content in a way other systems can identify. OpenAI said the watermark will roll out over the coming weeks to eligible ChatGPT and Codex users on all plans, but only in the EU. Developers using OpenAI’s API anywhere in the world can turn it on for select models starting today; it’s off by default. OpenAI said it is not making text watermarking a global default at launch. The watermark is not an actual symbol, but works by subtly shaping the model’s word choices, leaving a pattern readers can’t see, but a detector can pick up. Because it lives in the words themselves, it travels with the text when it’s copied and pasted. OpenAI said the watermark doesn’t identify the user, and that it saw no meaningful change in its models’ performance with it switched on. OpenAI also published a technical report for its method, called textGrain, alongside the announcement. Co-written with researchers from the University of Pennsylvania and Yale, it walks through an example of using a secret key to sort next-word predictions to finish the sentence. Add hundreds of these nudges together, and the detector can spot AI-generated content using only the text and the key. Can the watermark be removed by editing? OpenAI’s tests suggest yes. In one test, replacing 10% of words with synonyms dropped detection from about 92% to 66%. The company also said short passages, math answers, and translated text are harder to detect. “These limitations contribute to our decision to provide initial detector access only to approved researchers and expert organizations, who can help us evaluate reliability and responsible uses,” said the company. OpenAI also cautioned that a missing watermark “does not prove human authorship.” The text could be too short or too heavily edited, or it could come from another company’s AI. “[Watermarks] can indicate that an OpenAI system generated or processed part of a passage, but not how much human judgment, editing, or creativity went into it,” the company said. The announcement comes two months after Anthropic said it would watermark text generated by Claude, a move it’s applying worldwide. That decision drew backlash from some Claude users, who argued they had supplied “the instructions, context, decisions” while Claude was just “the tool.” OpenAI had built a text watermark before but held off on releasing it, partly over concerns that users would switch to rivals that didn’t watermark, The Wall Street Journal reported in 2024. Anthropic, Google, Meta, Microsoft and OpenAI are among the companies that have committed to following the EU’s code of practice on AI-generated content. Topics AI, eu, Government & Policy, OpenAI, TC When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence. Aditya Mehta Editorial Fellow Aditya Mehta is a reporter at TechCrunch covering AI. He’s supported by the Tarbell Center for AI Journalism and attended UC Berkeley. You can contact from Aditya by emailing aditya.mehta@techcrunch.com or via encrypted message at adymehta.74 on Signal. View Bio

发现内容有误?提交纠错