字号·· | 护眼
arstechnica

OpenAI将默认为ChatGPT输出添加水印,但仅限欧盟OpenAI will watermark ChatGPT outputs by default—but only in the EU

点「原文对照」整页切到原文,或双击某段只看那段的原文。

OpenAI宣布将开始在欧盟自动为ChatGPT生成的文本添加水印。该公司还将在其他地区提供水印功能,但在欧盟以外地区默认关闭。

OpenAI will begin automatically watermarking text generated with ChatGPT in the European Union, the company has announced. It will also offer the watermarking feature in other regions, but it will be off by default outside of the EU.

这一在欧洲的举措旨在遵守8月生效的《欧盟人工智能法》。该法要求以可供其他工具检测的方式标记人工智能模型生成的内容。遗憾的是,目前还没有完全有效且可靠的解决方案。已经存在一些标准,如SynthID和C2PA项目,但它们对于具有基本知识的人来说相对容易绕过。

The move in Europe is driven by a need to comply with the EU AI Act, which took effect in August. It requires marking content produced by AI models in a way that another tool can detect. Unfortunately, there is still no completely effective and reliable way to do that. A few standards already exist, like SynthID and the C2PA project, but they are relatively easy to circumvent for anyone with basic know-how.

OpenAI的水印也可能存在类似问题。其方法具有专有性,公司将其称为textGrain,并发布了一篇技术论文解释其工作原理。但总体来说,它的运作方式与我们过去见过的其他LLM水印工具类似:它在词汇选择中嵌入人类读者难以察觉的模式,这些模式不会显著改变输出的整体质量,但拥有密钥的人可以使用专门的检测器来发现它们。OpenAI表示,将向有限数量的研究人员和组织提供检测器的访问权限,并为其他人建立审批请求流程,以便逐步添加更多用户。

The same is likely true for OpenAI's watermark. Its method is proprietary; the company calls it textGrain, and has published a technical paper explaining how it works. But in general, it works like other LLM watermarking tools we've seen in the past: It puts patterns in the word choices that are not clear to a human reader, and that don't meaningfully change the general quality of the output, but that someone with a key can use a specialized detector to find. OpenAI says it will be giving access to the detector to a limited number of researchers and organizations, and providing a request-for-approval process for others to be added over time.