OpenAI to watermark ChatGPT outputs by default, only in the EU
openai regulation
| Source: Mastodon | Original article
OpenAI will begin automatically watermarking ChatGPT responses for users in the European Union, while the feature remains optional elsewhere.
OpenAI announced on Monday that it will embed an invisible watermark in every text response generated by ChatGPT – and Codex code snippets – for users located in the European Union. The move is a direct response to the EU AI Act, which obliges providers of high‑risk generative systems to attach machine‑readable provenance information to AI‑generated content.
The watermark is not a visible label; instead, it is a subtle pattern woven into the token sequence that can be detected by OpenAI’s own detection tool. The company admits the technique is “not especially reliable” and can be easily sidestepped. In internal tests, replacing just 25 % of the words with synonyms reduced detection rates to around 17 %, underscoring the fragility of the approach.
Why it matters is twofold. First, the watermark demonstrates OpenAI’s willingness to adapt its products to meet the bloc’s regulatory demands, a step that could set a precedent for other AI firms facing similar obligations. Second, the limited robustness of the watermark raises questions about the practical enforceability of provenance rules and whether they will meaningfully curb the spread of undisclosed AI‑generated text.
What to watch next includes OpenAI’s rollout timeline and whether the watermark will be extended beyond the EU, especially as other jurisdictions contemplate comparable labeling requirements. Regulators may also test the detection tool’s efficacy in real‑world settings, potentially prompting tighter standards or new technical solutions. Finally, developers and users will be monitoring any impact on content quality or workflow, as well as the emergence of third‑party tools designed to strip or spoof the watermark.
Sources
Back to AIPULSEN