PaPoo
cover

Claude Code Is Not Trying to Replace Your Editor, and That’s the Point

What surprised me here is how unapologetically boring the answer is. “Use AI for everything” is a seductive slogan, but Boris Cherny’s position sounds more like engineering discipline than product hype: let Claude Code be the thing that edits, checks, and reasons over code, but don’t pretend it should swallow the whole workflow. I think that’s the healthier answer, even if it’s less exciting.

The bit that actually lands is the “blockbox” idea — not in the sense of secrecy, but as a boundary. If you let an agent roam across the entire codebase, it will happily make changes you didn’t really ask for, or fix the wrong layer because it can. Cherny’s framing is basically: keep the model inside a box where its output can be verified. That sounds obvious, but a lot of teams still use agentic tooling as if trust is free. It isn’t.

I also like that the article doesn’t romanticize “trust the AI.” The practical guards listed here — lint, unit tests, e2e tests, fuzzing, code review, keeping specs current — are not glamorous, but they’re the difference between “wow, it wrote a lot” and “we can actually ship this.” The line about holding the bar on code quality is the right one. If Claude can’t meet the bar, the answer isn’t to lower the bar.

Where I’m a little less convinced is the implied confidence that this can be cleanly reduced to three steps. Real projects are messier than that. A lint pass and a few tests are great, but they don’t catch architectural weirdness, performance regressions, or the subtle ways agents can optimize for the wrong thing. The article gestures at that with “scope,” and that’s probably the real crux: the smaller and more legible the task, the more useful Claude Code gets. The broader the task, the more you need human judgment in the loop.

What I’d actually try, if I were using Claude Code seriously, is exactly the workflow the article hints at: give it tightly bounded changes, write the tests first or at least alongside the change, and make sure the repo has clear project rules in CLAUDE.md. If the team can’t describe the work in a way an agent can safely execute, that’s usually a sign the work itself is under-specified, not a sign the model is too dumb.

That’s the part I think some people will miss. The article is not really saying “agents are limited.” It’s saying “good engineering is still required.” Claude Code doesn’t remove the need for craftsmanship; it raises the cost of pretending you can skip it.


Reference: AI is not something to “black-box”? Claude Code’s creator explains the prototype type and the reality of code generation

同じ著者の記事