OpenAI 將在歐盟預設為 ChatGPT 生成文字加上浮水印,其他地區功能預設關閉
OpenAI will watermark ChatGPT outputs by default—but only in the EU
OpenAI 宣佈,將在歐盟預設為 ChatGPT 生成的文字加上浮水印,其他地區亦可使用此功能,但預設關閉。此舉是為遵守 8 月生效的 EU AI Act;其專有方法 textGrain 會在用詞選擇中加入人眼不易察覺、且不會明顯影響輸出整體質素的模式,供持有密鑰者以專用偵測器識別。SynthID 和 C2PA 等既有方案相對容易被具基本知識者繞過,OpenAI 的浮水印也可能有相同限制。
OpenAI will begin automatically watermarking text generated with ChatGPT in the European Union, the company has announced. It will also offer the watermarking feature in other regions, but it will be off by default outside of the EU.
The move in Europe is driven by a need to comply with the EU AI Act, which took effect in August. It requires marking content produced by AI models in a way that another tool can detect. Unfortunately, there is still no completely effective and reliable way to do that. A few standards already exist, like SynthID and the C2PA project, but they are relatively easy to circumvent for anyone with basic know-how.
The same is likely true for OpenAI's watermark. Its method is proprietary; the company calls it textGrain, and has published a technical paper explaining how it works. But in general, it works like other LLM watermarking tools we've seen in the past: It puts patterns in the word choices that are not clear to a human reader, and that don't meaningfully change the general quality of the output, but that someone with a key can use a specialized detector to find. OpenAI says it will be giving access to the detector to a limited number of researchers and organizations, and providing a request-for-approval process for others to be added over time.
來源:Ars Technica · AI · arstechnica.com