New models will mark AI-generated content from day one: Claude will now hide an invisible watermark inside ordinary words heres how thats even possible, and how EU rules could push OpenAI and Google to follow suit
Date:
Wed, 12 Aug 2026 13:00:00 +0000
Description:
Anthropic has announced that new Claude models will hide an invisible watermark inside any generated text heres how thats even possible, and how OpenAI and Google might have to follow suit.
FULL STORY ======================================================================Copy link Facebook X Whatsapp Reddit Pinterest Flipboard Threads Email Share this article 0 Join the conversation Follow us Add us as a preferred source on Google Newsletter Subscribe to our newsletter To comply with the EU AI Act's Article 50(2) Code of Practice on Transparency of AI-Generated Content, Anthropic has announced that New models will mark AI-generated content from day one.
This is a remarkable step for Claude , because it not only applies to any images it generates, which are relatively easy to watermark, but also to any text it generates. Anthropic says that Claude models will have an imperceptible watermark embedded directly into generated text at the model level. It says the mark should survive copying/pasting and minor edits.
Latest Videos From TechRadar Watch full video here:
Anthropic has not yet publicly explained the exact algorithm or released a detector, so we can only speculate for now about how it might be doing this and its effectiveness. Hemorrhaging customers Personally, I think that Claude is about to start hemorrhaging customers, unless the marking is relatively easy to circumvent, or all the other major AI players immediately follow
suit. You may like Anthropic's allegation that Alibaba copied Claude has huge implications After saying Mythos was too dangerous, Anthropic just launched a public version Claude Sonnet 5 is here, and it's the 'most agentic Sonnet model yet'
If every piece of text it produces will now be easily identified as AI, then Claude becomes useless as a tool to a lot of people who are currently using
it to generate text and are not being entirely honest about where that text came from.
And then theres the issue of using Claude to proofread your human-written
text will your text now be flagged as AI if you accept Claudes editing advice? Get daily insight, inspiration and deals in your inbox Sign up for breaking news, reviews, opinion, top tech deals, and more. Contact me with news and offers from other Future brands Receive email from us on behalf of our trusted partners or sponsors By submitting your information you agree to the Terms & Conditions and Privacy Policy and are aged 16 or over.
Your initial reaction to that might be Good! You should be forced to reveal when AI has written something, and its about time people started writing on their own again!, and youd be entirely justified in that opinion.
But while it remains the only one of the big three AIs thats doing this, I think well see a lot of people switch to either ChatGPT or Gemini, because they dont want to be revealed as using AI in their work. (Image credit: Shutterstock/ gguy) How is watermarking plain text even possible? We dont
know exactly how Anthropic is doing its marking with text yet, but my best guess is statistical watermarking during token generation rather than hidden Unicode characters or metadata. Imagine that at every point Claude is
choosing among several perfectly reasonable next words: What to read next Alibaba is banning its workers from using Claude Code as US v China AI battle heats up 'Bringing Claude Tag into Slack is about making AI multiplayer': You can now tag Claude directly in Slack Claude Sonnet 5 booked nothing, and
still felt like an assistant
e.g. The movie was excellent / superb / terrific / impressive .
Normally it chooses according to the model's probability distribution. A watermarking system can secretly divide possible tokens into preferred and non-preferred groups using a key. Claude then gives a tiny statistical nudge toward the preferred group.
One word tells you nothing. But across 500 or 1,000 words, a detector with
the key can ask if the text is choosing the preferred tokens significantly more often than chance would allow. If yes, then there's statistical evidence it came from the watermarked model.
So, the watermark is more like a faint statistical fingerprint distributed across hundreds of choices, which would also explain how it can survive a copy-and-paste. Youre essentially copying the fingerprint along with the words. But can you crack the code? Since Anthropic hasnt released an AI text detector yet, its impossible to know how easy this code will be to break just by changing a few words. For instance, if you put your 1,000-word Claude article into another LLM and wrote Rewrite this completely in different
words while preserving the meaning , would it then be impossible to detect as AI?
Id also be interested to know how long a piece of text has to be before it
can be marked in this way, and as soon as a detector is made available, Ill
be testing it.
Perhaps the bigger issue is that Claude has done this to comply with Article 50(2) of the EU AI Act. From August 2, 2026, providers of generative AI systems that produce text, images, audio, or video are required to make those outputs machine-readable and detectable as artificially generated or manipulated, insofar as that is technically feasible. Existing systems get a limited transition period until December 2, 2026 for this particular requirement.
A note on Anthropics statement confirms that the watermarking additions will be applied retroactively to all existing Claude models, not just any new models it produces.
Anthropic's new system is explicitly a response to those rules, and it says the watermark will apply globally, not just when Claude is being used in Europe. Broadly speaking, OpenAI and Google face the same requirement if they want to offer qualifying generative-AI systems in the EU.
They don't necessarily have to copy Anthropic's method of using statistical text watermarking, but they will need to produce output that is machine-readable and detectable, and the technical solution should be effective, interoperable, robust and reliable as far as technically feasible.
Ive contacted OpenAI and Google for comment, and will update this article if
I receive it. For now, I think if Anthropic embarks on this path as the only one of the major three AI chatbots to do so, it could have a disastrous
effect on its customer base. Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.
======================================================================
Link to news story:
https://www.techradar.com/ai-platforms-assistants/claude/new-models-will-mark- ai-generated-content-from-day-one-claude-will-now-hide-an-invisible-watermark- inside-ordinary-words-heres-how-thats-even-possible-and-how-eu-rules-could-pus h-openai-and-google-to-follow-suit
--- Mystic BBS v1.12 A49 (Linux/64)
* Origin: tqwNet Technology News (1337:1/100)