PaPoo
cover

Cheap Haiku, or just cheaper Claude?

What jumped out at me isn’t the price cut by itself. It’s Anthropic’s insistence that the smallest model in the family is now something you should use for agentic work. That’s a meaningful push, because “small and cheap” usually means “good enough for drafts, not for loops.” If Haiku 5.5 really moved the needle on computer use and command-line tasks, then Anthropic is trying to make the case that the low end isn’t a toy anymore.

I’m still a little skeptical of benchmark-heavy claims here. The article leans on gains in computer-use and CLI tests, which are exactly the kind of numbers that can look exciting while hiding brittleness in real workflows. Agentic stuff lives or dies on the messy bits: login flows, weird shell output, half-broken web apps, and the long tail of task failures. A model can score well and still feel annoying in practice. I’d want to see how often Haiku 5.5 actually completes a multi-step task without needing babysitting.

The “effort controls” detail is the most interesting part to me. That sounds like Anthropic is giving developers a way to trade speed, cost, and maybe quality more deliberately instead of treating model choice as a blunt instrument. If that’s the direction, it makes sense. A lot of agent workloads do not need the biggest model all the time; they need a cheap model that can be told when to think harder. That’s a more honest product story than pretending every problem deserves the flagship.

What I’d try first is boring but revealing: run Haiku 5.5 on a real internal support or ops workflow, not a benchmark demo. See whether the cheaper model still stays coherent across tool calls, retries, and failure recovery. If it does, then this is more than a pricing story. If it doesn’t, then it’s just Anthropic making the small model look better on paper.


Reference: Anthropic launches Haiku 5.5 at a much lower price

同じ著者の記事