PaPoo
cover

Claude is doing more of Anthropic’s research than I expected

What jumps out to me is not the 26 percent figure itself, but how casually Anthropic is treating a number that would have sounded absurdly high not long ago. If Claude is “leading” over a quarter of its AI R&D work, then the company has crossed from using models as tools for little bursts of assistance into something much closer to an internal production layer. That’s a real shift. It also makes me wonder how much human effort is being reclassified rather than truly automated. “Leads” is doing a lot of work here.

I’m also not fully convinced by the framing around safety transparency. Anthropic is proposing measurement standards, which is good, but measurement alone doesn’t slow anything down. It can just as easily become a way to make acceleration feel legible and therefore manageable. Maybe that’s useful. Maybe it’s also a bit of a fig leaf if the underlying incentive structure stays the same. A chart showing “automation level” since August 2025 sounds impressive, but I’d want to know how hard it is to game, how often the third-party validation actually bites, and whether the measure captures the messy parts of research work that models still fumble.

The more interesting claim, to me, is the one about Claude doing “large chunks of work under close human direction” on more than 90 percent of their research. That sounds plausible, and also exactly like the sort of statement that can hide a lot of ambiguity. Large chunks of what? Literature review, code drafting, experiment scaffolding, analysis, internal docs? Those are very different tasks. If I were building with Claude, I’d care less about the headline percentage and more about where the model is genuinely saving time versus where it is just shifting the human’s job from doing to checking.

Anthropic seems to be nudging the industry toward a world where “AI-led R&D” becomes normal jargon. I think that’s the real story. Not the number, but the fact that a frontier model company is already talking as if its own research loop is partly inside the machine. That is both unsettling and, if you’re building with these tools, kind of inevitable.


Reference: Anthropic says Claude 'leads' 26 percent of its AI R&D work

同じ著者の記事