PaPoo
cover

Anthropic’s watermark feels broader than the law needs

What jumps out to me is not the watermark itself, but how quickly a compliance story turns into product policy. Anthropic seems to be treating “can be detected later” as the default for anything Claude touches, even light editing. That may be defensible as a safety posture, but it also feels like the kind of move that quietly expands regulatory intent until it covers normal use cases that most people would not think of as AI-generated content.

I’m a little unconvinced by the “EU made us do it” framing. If the article is right, Article 50(2) is aimed at marking AI outputs when the system substantially changes the user’s words or is doing synthetic generation, not when it’s just polishing prose. If that’s the line, then tagging translated text, summaries, or heavily edited drafts starts to look like Anthropic making a broader judgment than the law itself requires. Maybe that’s the safest interpretation from a legal-risk standpoint. But safest for Anthropic is not the same thing as clearest for users.

That distinction matters because this kind of watermark is not a neutral little label. If the signal follows text through copy-paste and survives edits, it can easily be read as “Claude wrote this,” even when Claude only cleaned up a paragraph or fixed grammar. For people using LLMs as writing tools rather than ghostwriters, that’s a pretty important difference. I’d want to know much more about false positives, how often the signal survives ordinary human editing, and what kind of downstream systems might treat the mark as evidence of misconduct. The article quotes someone warning against exactly that, and I think they’re right.

There’s also a strategic angle here that feels more interesting than the legal one. Anthropic gets to wear the safety-first badge, while also pushing the burden of interpretation onto employers, platforms, and whoever is unlucky enough to receive the text. That’s smart in a corporate sense. It is not obviously kind to users. If I were building on Claude, I’d be asking a blunt question: is this a provenance hint, or is it going to become an accusation machine once it lands in enterprise tooling?

The most telling part is that other companies seem to be heading toward something similar, just with different levels of transparency and specificity. That means this probably won’t stay an Anthropic quirk for long. The real fight is over where the line sits between “AI assisted” and “AI authored,” because the industry keeps wanting that line to be machine-readable even when the work itself is messy and collaborative. I think that’s the part everyone should be wary of.


Reference: Why Anthropic's AI watermark for Claude text goes further than rivals — for now

同じ著者の記事