The continuum Band 03
The agent plans, edits, runs tests and iterates. The human sets goals and reviews.
Also called: Agentic engineering, coding agents
A coding agent takes a scoped task, reads the codebase, makes multi-file changes, runs commands and tests, and iterates until it believes the task is done. The human sets the goal, supplies context and conventions, approves risky actions and reviews the diff before it merges.
Karpathy later preferred "agentic engineering" as the framing for serious work, as distinct from vibe coding.
Agents that could edit code and run tests appeared in research first, including Princeton’s SWE-agent and Cognition’s Devin in 2024. Terminal-based agents such as Claude Code, OpenAI Codex and Gemini CLI followed in 2025 and brought the approach into everyday use.
Karpathy proposed “agentic engineering” in February 2026 and set out the idea at Sequoia’s AI Ascent in April. Others, including Simon Willison and Addy Osmani, had used the phrase earlier.
Steps 2–4 repeat inside one task.
Typical tools: Claude Code, OpenAI Codex, Gemini CLI, Cursor agent mode, GitHub Copilot coding agent.
Solid: AI does core work · Hatched: partial or informal · Dashed: done by people. Compare all bands
Good for: Multi-step internal features with a mostly clear goal
Where it runs: A terminal or editor agent (Claude Code, Codex, Cursor) working in a repository with tests.