Using AI to crank out everything imaginable, from your homework to emails to code, is great and all — until you have to own up to it.
Fearless embracers of AI are suddenly clutching at their pearls, after Anthropic announced that its Claude chatbot will watermark the text it generates, potentially exposing anyone who wants to get away with using the tech without detection.
The announcement has caused a meltdown in AI circles.
“This watermark is the dumbest f*cking thing I’ve ever heard in my life. Are they going to ask you to provide an ID so you can write non-watermarked text?” seethed one user on r/ClaudeAI.
“That mark will be the kiss of death on any piece of text that people can sell,” another said. “People won’t want to pay for it. It will be a scarlet letter.”
There were occasional injections of levity.
“I thought it was already watermarking text by including an em dash every 4 words,” one joked.
But TechCrunch spotted an especially dramatic breakdown on the r/artificial subreddit, where a user who goes by visionode posted an extended rant comparing the watermarking scheme to nefarious police tactics and to systemic oppression. No, really.
“Who will get caught? You. The student who used Claude to reorganize a paragraph. The journalist who asked the AI to summarize a two-hundred-page transcript. The writer who had creative block and asked for synonyms,” visionode wrote. “Those guys come out of the process with a digital tattoo on their forehead.”
“You know what this reminds me of?” visionode asks. “Those police operations that arrest the drug user and leave the dealer alone. Watermarking is the same thing.”
“And the stigma,” the user continued. “We’re creating a caste of ‘dirty’ creators. People who dared to use a tool.”
Unfortunately for visionode, not many could get behind their heavily-downvoted post.
“Did you generate this histrionic screed with a chatbot too?” one replier asked.
Anthropic said it was implementing the watermark system in response to the European Union’s landmark AI Act passed in 2024, which requires that AI companies mark content that’s been generated or edited by their systems. It works by making subtle changes in the AI’s word choices across the text it generates, which are supposed to be imperceptible to a human but, in aggregate, form a pattern that is detectable with the tool.
It’s definitely not a bulletproof approach. Anthropic says that the watermarks will “persist through some editing,” but if it’s pasted and rewritten with another chatbot, that signal could be destroyed, Ars Technica noted in its breakdown. And there’s a worry that the word choices the AI goes with to create a watermark might deteriorate the quality of its prose. Terrifyingly, AI users may have to start polishing their writing without a chatbot.
Worst of all, Ars warns, once Anthropic releases how its detection tool works, it’ll be easy for bad actors to create a tool that goes in and erases the watermarks. We’re already starting to see that happen with SynthID, Google DeepMind’s own system for embedding hidden telltales of AI provenance in images — though no one has figured out how to fully remove its watermarks yet.
More on AI: Game Dev CEO Accused of Replacing Writers With AI is Now Completely Crashing Out