Agentic coding tools
Claude Code, OpenAI Codex, Cursor, GitHub Copilot, Cognition's Devin Desktop (formerly Windsurf)
Where they are genuinely better
Genuinely excellent, and we say so. Repo-wide reasoning, terminal-native execution, tight feedback loops, and productivity gains that show up in shipped software. If you have engineers, buy these. We use them ourselves.
Where they stop
They are aimed at work that has an oracle — the tests pass or they don't — and a shared professional idiom. Their steering layer has grown up: rules files, session memory, organisation-level custom instructions. But that captures a team's conventions, not how your specific person decides. And none of them, in their current public documentation, grades output against a named human expert's withheld judgment; what they ship is rubric-based model-as-judge evaluation and automated code review.
Complement, never competitor. Keep them. They don't do this and they aren't trying to.