Texts written by AI, because the new controls can be circumvented
OpenAI is also starting to attribute the texts produced by its artificial intelligence to distinguish them from those written by humans.
OpenAI is also beginning to watermark the texts generated by its artificial intelligence to distinguish them from those written by humans. Following Anthropic, which announced a watermark for Claude in August, OpenAI announced on 5 October that, in the coming weeks, it will introduce an invisible mark in texts produced by ChatGPT and Codex within the European Union. However, the company’s own published findings show just how fragile this traceability is: changing just a few words can be enough to allow much of the content examined to slip through the net.
For businesses, schools and publishers, the prospect is an attractive one: to have an indication of a document’s origin, without relying solely on the statements of the person submitting it. The risk, however, is attributing a value to that indication that it does not possess. One text may have been produced entirely by a machine and passed the check; another may have been created by a person but retain traces of subsequent automated reworking.
The impetus comes from the AI Act. The European transparency provisions, which came into force on 2 August, require providers to ensure that synthetic content is recognisable by automated tools, with certain exceptions, including specific functions designed solely to assist with editing. For systems already on the market prior to that date, the transitional deadline for labelling is 2 December 2026. The obligation creates a shared incentive to develop these tools, but does not eliminate their technical weaknesses.
A text watermark is not a hidden label within the document. It does not consist of invisible characters that can be deleted, nor of information that disappears when a passage is copied into another programme. As illustrated by the study on SynthID-Text published in *Nature* in 2024, the watermarking takes place during generation: from among the various plausible continuations of a sentence, the system makes its choice according to a mechanism that leaves a recognisable statistical pattern.
The reader sees a normal sequence of words. The detector, on the other hand, looks for that pattern across the entire text. This is a crucial distinction: changing the font or pasting the text without formatting does not alter the words and therefore does not, in itself, remove the signature.

