Topic · Concept
Coding agents
Coding agents are AI systems that take a software task, such as fixing a bug or adding a feature, and carry it out over many steps: reading code, editing files, running commands and tests, and returning a change for review. This page covers how they work, how they are measured and where they fail.
Fact file
Canonical agent vs workflow definitionAnthropic, "Building effective agents," December 19, 2024
SWE-bench VerifiedOpenAI stopped evaluating on it (Feb 23, 2026), citing flawed tasks and contamination; recommends SWE-bench Pro
Claude Code default modeAuto mode is the built-in starting permission mode for interactive terminal and VS Code sessions from v2.1.283
Example productsClaude Code (GA May 22, 2025), Codex (cloud agent May 16, 2025), Cursor, GitHub Copilot
Shared instruction formatAGENTS.md, a founding project of the Agentic AI Foundation (Dec 9, 2025)
Latest on Coding agents
Oct 9Agents · ExplainerClaude Code mods, explained: what Anthropic's new extension layer can do, where it runs, and the security trade-offOct 9Developer tools · GuideBuild your first agent with the Claude Agent SDK (TypeScript, October 2026)Oct 9Developer tools · GuideHow to write a CLAUDE.md or AGENTS.md file that coding agents actually followOct 9Agents · ExplainerWhat "agentic" really means in 2026Oct 9Agents · ExplainerWhat is MCP? The Model Context Protocol explained (2026 edition)
Nothing in this format on this page — try another filter or the next page.
About this hub
Coding agents are AI systems that take a software task, such as fixing a bug or adding a feature, and carry it out…