OpenAI will watermark ChatGPT outputs by default—but only in the EU
摘要
OpenAI宣布将在欧盟默认对ChatGPT生成的文本添加水印,其他地区可选择启用但默认关闭。此举旨在遵守8月生效的欧盟《人工智能法案》,该法案要求以可被工具检测的方式标记AI生成内容。OpenAI的水印技术名为textGrain,通过调整用词模式嵌入人类不易察觉、需专用检测器识别的标记,目前仅向有限研究者和机构开放检测权限。
OpenAI will begin automatically watermarking text generated with ChatGPT in the European Union, the company has announced. It will also offer the watermarking feature in other regions, but it will be off by default outside of the EU.
The move in Europe is driven by a need to comply with the EU AI Act, which took effect in August. It requires marking content produced by AI models in a way that another tool can detect. Unfortunately, there is still no completely effective and reliable way to do that. A few standards already exist, like SynthID and the C2PA project, but they are relatively easy to circumvent for anyone with basic know-how.
The same is likely true for OpenAI's watermark. Its method is proprietary; the company calls it textGrain, and has published a technical paper explaining how it works. But in general, it works like other LLM watermarking tools we've seen in the past: It puts patterns in the word choices that are not clear to a human reader, and that don't meaningfully change the general quality of the output, but that someone with a key can use a specialized detector to find. OpenAI says it will be giving access to the detector to a limited number of researchers and organizations, and providing a request-for-approval process for others to be added over time.
转载信息
评论 (0)
暂无评论,来留下第一条评论吧