AI-generated text now carries an invisible watermark. What is it good for?
Translated from Dutch and summarized by DistantNews. Read the original for the full story.
At a glance
- European rules now require AI companies to watermark text and images generated by their models to help combat misinformation and manipulation.
- Anthropicโs system subtly influences the random selection of otherwise suitable words, creating a detectable pattern without making the watermark visible to readers.
- Researchers and an AI expert say the method is designed to preserve variation and text quality, while OpenAI is also working on watermarking ChatGPT text.
AI companies must now add watermarks to text and images generated by their models under a European requirement introduced last month. The aim is to counter misinformation and manipulation, but the watermark remains invisible to readers.
Anthropic, the US company behind the Claude chatbot, has explained its approach in a detailed blog post. The technology turns out to be a version of SynthID, a system developed by Google DeepMind researchers in 2024. Googleโs Gemini chatbot has added a watermark since November, while OpenAI says it is working on a similar system for ChatGPT text.
The watermark lies in the modelโs choice of words. A large language model predicts what word should follow the words already generated, taking the prompt and many other variables into account. Often, several words or phrasings would work equally well. A chatbot can therefore produce different suitable answers to the same question.
SynthID uses that built-in randomness. If โcloudyโ and โgreyโ both fit the sentence โthe sky is...โ, the model assigns each option a random score and chooses according to probability. Without a watermark, the scoring is entirely random. With SynthID, a secret key slightly influences the scores, giving one option a small statistical advantage. The preferred word does not become mandatory, so the text remains varied.
โThis is really cleverly done, because this watermarking method mathematically ensures that the quality of the output does not decline,โ said Rik van Noord, an AI researcher and large-language-model expert at the University of Groningen. All selected words remain within the modelโs list of options considered suitable for the context, he said. Google DeepMind researchers reached the same conclusion after analyzing 20 million texts generated by Gemini and comparing them with other material.
This is really cleverly done, because this watermarking method mathematically ensures that the quality of the output does not decline.
Originally published by NRC Handelsblad in Dutch. Translated, summarized, and contextualized automatically by DistantNews, with a note on how the source frames the story. Not individually reviewed before publishing. How this works.