Identification markers in text as well as images

Global artificial intelligence (AI) companies are successively introducing watermarking technology that can identify text written using generative AI. Just as markers are embedded in AI-generated images to indicate that they are AI-created, companies are looking to apply similar watermarks to text. As the European Union (EU), South Korea and jurisdictions around the world have legislated the use of watermarks for AI-generated content, the scope of their use is expected to expand gradually.


According to the AI industry on October 9, OpenAI announced on October 5 local time that it would introduce Textgrain, a technology that identifies whether text was generated by AI. OpenAI plans to apply the technology to ChatGPT and Codex in Europe within the next few weeks.


"Now They'll Catch Every AI-Written Personal Statement"... A 'Secret Fingerprint' Embedded in Sentences [Into the AI World] View original image

Textgrain embeds subtle statistical signals as an AI model selects words while generating text. By applying specific statistical rules when choosing words, it creates a kind of "pattern" throughout the text that can indicate whether AI was used. Because the method does not insert hidden characters or special symbols into sentences, the text can still be identified even if it is copied and pasted elsewhere.


OpenAI said that applying the technology would make no material difference to the quality, performance or readability of AI responses. However, because of how Textgrain works, its detection rate may be lower for short passages or text that uses few words. The detection rate also falls if users change some of the words or expressions in AI-generated text.


OpenAI is also accepting requests for access and will initially provide a tool for detecting Textgrain to approved researchers and professional institutions. Those given access to the tool will be able to determine whether text was generated by AI.


Global AI leaders, including Anthropic and Google, have already introduced similar text watermarking technology. Anthropic introduced a text watermarking feature using a similar method in August. Google has also applied SynthID-Text, a text watermarking technology developed by Google DeepMind, to text generated through Gemini since May 2024.


The EU's AI Act is behind the adoption of text watermarking technology by global companies. Similar to South Korea's AI Basic Act, the EU law includes a provision requiring transparency for generative AI. It requires technically detectable watermarks or identifying markers in text as well as media content such as images, video and audio.



The measures were introduced as AI-generated content becomes more sophisticated, increasing the risk that it could be exploited for deepfakes, cybercrime and the spread of false or manipulated information. With concern about AI safety also growing amid calls to slow the pace of development in response to recent advances in AI models, the prevailing view is that the use of watermarking technology will expand.


This content was produced with the assistance of AI translation services.

© The Asia Business Daily. All rights reserved. Unauthorized AI training and use prohibited.

Today’s Briefing