The Claude watermarking witch hunts are coming

by | Aug 15, 2026

A central black-and-white portrait is split in two: one half remains intact while the other dissolves into torn text fragments and marks inside a distressed red circle. Shadowy crowds with pitchforks and torches rise from both lower corners, while a dotted cut line and scissors suggest the watermark can be removed.

Claude’s outputs are going to be watermarked. When a detector becomes available to the public, the witch hunts are gonna get INTENSE!

Innocent people are going to get buckets of water thrown over them, and the real wicked witches will be laughing hysterically from their broomsticks.  

Seriously, if you thought the hysteria over em dashes was bad, just imagine the pitchforks and torches we’ll see brandished when a detector indicates that Claude was involved. 

And it’s not just Claude either. Most of the major AI labs will be watermarking. 

Oh, the angry mobs are coming alright. I can hear them chanting “AI SLOP! AI SLOP!”

A clipped stack of handwritten pages is torn open along one side, revealing a dense neural-network-like structure of nodes, connecting lines, mathematical symbols and algorithmic fragments beneath the paper. A weathered red geometric block crosses the boundary between human writing and the exposed computational layer.

The problem with watermarking

In simple terms, watermarking works by algorithmically influencing an AI’s word choice so the patterns can be detected. 

There are two problems with using a watermark to identify “AI Slop”:

1. Editing the output enough to disrupt the pattern means the watermark is effectively removed.

2. Using AI to edit your own original work could be enough to embed the watermark. 

So the watermark doesn’t prove that AI was the author, because it might have edited a legitimately human-authored work.

And the absence of a watermark doesn’t prove a human-authored the piece because the watermark could have been removed. 

The real wicked witches will use technical methods to easily automate the removal of the watermark.

Meanwhile, the pitchforks could be out for people who are simply using AI to edit their own writing.

A sparse network of isolated profile silhouettes and faint connecting lines spreads across the left, contrasted with a tightly packed swarm of profiles, red fragments and connections contained inside a large distressed red circle on the right. A single profile at the lower right is separately circled, suggesting the danger of targeting an individual instead of recognising the larger pattern.

So is watermarking useless?

Watermarking isn’t entirely useless.

It could be very useful when used as one signal for detecting AI misuse at scale. 

For example, think about Facebook being able to identify that the proportion of AI-generated content normally sits at, say, 35%.

Then, an election is coming up, and they see a spike of up to 85% in a series of new accounts. That would be one signal that might allow them to more quickly identify that those accounts should be checked for misinformation. 

Yeah, ok, I know, it’s a stretch to think of Facebook using the tools to actually improve the quality of content on their platform. But hypothetically, it’s possible. 

That’s how I think watermarking should be used. 

Not to shame individuals who use AI, but to make the internet a better place by identifying large-scale patterns of use. 

Nonetheless, I think the witch hunts are coming. 

Individuals will be burned at the stake because their content was watermarked. 

A lone sculptor chisels a rough mass of fragmented symbols, network lines and typographic debris into a clean rectangular slab. The finished surface reads “Thought. Selection. Structure. Clarity. Meaning. — Human.” with a muted red circle below, representing human judgement and craft shaping algorithmic material into coherent work.

AI use itself is not the problem

I’m not just defending the people who use AI for light editing. I believe it’s entirely possible to use AI to create meaningful and engaging content. 

It takes a lot of work.

It requires human thought, curation, and guidance. 

It usually requires significant editing or directing of AI outputs. 

But a mindful human operator can absolutely create good-quality writing with AI.

The witch hunts are coming. But eventually the mob will exhaust themselves. At some point we’ll learn to burn people at the stake only when they actually cast evil spells and cause harm.

Not when they use an em dash.

Frank Prendergast

Frank Prendergast

I've over two decades of experience helping businesses with their online presence. I'm also the owner of the most-talked-about moustache in the marketing world and I'm the Frank half of the award-winning digital marketing team Frank and Marci. Follow on LinkedIn