Head to head · Agent harness · October 2026 research run
Claude Code vs Devin
Claude Code scores 61.9 (C) on agent readiness against Devin's 55.7 (C), and leads in 5 of 7 scored categories. Both do agent harness.
Which one, for what
Good for A developer or a pipeline that wants Claude doing repository work with fine-grained, centrally managed permissions and an SDK for the same loop.
Ahead on
- Reliability, 50 against 30
- Agent ergonomics, 87 against 62
- Security & auth, 80 against 72
- Maintenance & community, 82 against 63
- Transparency & trust, 81 against 69
Watch for
The sandbox is off by default and native Windows has none
Devin C
Good for A team that wants to hand whole tasks to a cloud agent and collect pull requests, driven from a pipeline or another agent.
Also in its favour
- A hosted endpoint, with nothing to install
- No incidents deducted, where Claude Code loses 6 points for them
Watch for
www.devinstatus.com lists three critical incidents (22 July, 13 August, 24 September 2026) and six major ones on the cloud agent or web app since 10 July 2026
Score by category
| Category | Weight this run | Claude Code | Devin | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 50 | 30 | Claude Code +20 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 80 | 80 | even |
| Agent ergonomics | 13%16.2 | 87 | 62 | Claude Code +25 |
| Security & auth | 14%17.5 | 80 | 72 | Claude Code +8 |
| Payments & pricing | 10%12.5 | 20 | 20 | even |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 82 | 63 | Claude Code +19 |
| Transparency & trust | 7%8.8 | 81 | 69 | Claude Code +12 |
| Negative events | ≤15 | -6 | 0 | |
| Total | 61.9 · C | 55.7 · C |
Facts side by side
| Fact | Claude Code | Devin |
|---|---|---|
| Kind | Agent harness | Agent harness |
| Vendor | Anthropic | Cognition AI, Inc. |
| Hosted endpoint | no (local only) | https://mcp.devin.ai/mcp |
| Transports | HTTP | |
| Auth | OAuth or key | API key |
| Pricing | Paid | Freemium |
| x402 | no | no |
| Licence | Proprietary. LICENSE.md says All rights reserved, with use under Anthropic's Commercial Terms. The GitHub repository holds the changelog, plugins and examples, not the source | Proprietary service under Cognition's Platform Terms of Service |
| Tools exposed | none | 13 |
| Read-only variant documented | no | yes |
| llms.txt | yes | yes |
| Last release | 2026-10-01 | 2026-10-07 |
| Terms last updated | no date given | 2026-06-30 |
| Privacy policy last updated | no date given | 2026-03-09 |
| Customer content may train models | yes, with an opt-out | yes, with an opt-out |
| Terms restrict automated access | not found in the text | not found in the text |
| Terms restrict benchmarking | yes | yes |
| Terms or service can change without notice | not found in the text | not found in the text |
| Arbitration or class-action waiver | yes | yes |
| Popularity | 141k stars | none |
Verdicts
Claude Code
Six permission modes, allow, ask and deny rules down to command arguments, PreToolUse hooks, and managed settings that can disable bypass and auto mode. The sandbox is off by default and native Windows has none.
Devin
The v3 API is well specified, with a public OpenAPI 3.1 file covering 239 operations, problem+json errors, cursor pagination and service-user keys tied to roles. The status page records three critical and several major incidents on the cloud agent between 22 July and 24 September 2026, and no rate limit figures or retry guidance were found in the reviewed documentation.
Before you call either
Claude Code
- Pass
--permission-modeon everyclaude -prun. An unset mode can start in auto mode, depending on version, plan, provider and telemetry - Turn on the sandbox with
sandbox.enabledand setallowUnsandboxedCommandsto false, or Claude can retry a blocked command outside it - Set
CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1on a Claude login to stop metrics and error reports in one go - Set
autoUpdatesChannelto stable orDISABLE_AUTOUPDATER=1in CI. Native installs update themselves about once a day - Cap pipeline runs with
--max-turnsand--max-budget-usd
Devin
- Use a
cog_service-user key with the Member role. Legacyapk_keys fail against v3 and the MCP server with 401 or 403 - Set
max_acu_limiton every session you create. Usage is metered by the work done and has no published unit rate - Session creation is not idempotent in v3. After a timeout, list sessions by tag before creating again
- Enterprise keys and personal access tokens must send
X-Org-Idto the MCP server. Organisation-scoped keys resolve it automatically - Create scheduled work as an automation. POST to the schedules endpoint returns 403 for migrated organisations since 24 September 2026
Questions
Which is better for AI agents, Claude Code or Devin?
Claude Code scores 61.9 (C) on agent readiness against Devin's 55.7 (C), and leads in 5 of 7 scored categories.
Other comparisons with Claude Code or Devin
- Aider vs Claude Code
- Aider vs Devin
- Amp vs Claude Code
- Amp vs Devin
- Claude Code vs Cline
- Claude Code vs Cursor CLI
- Claude Code vs Pi
- Claude Code vs Gemini CLI
- Claude Code vs GitHub Copilot CLI
- Claude Code vs goose
- Claude Code vs Kiro CLI
- Claude Code vs OpenAI Codex
- Claude Code vs OpenCode
- Claude Code vs OpenHands
- Claude Code vs Prime Agent
- Claude Code vs Qwen Code
- Cline vs Devin
- Cursor CLI vs Devin
- Devin vs Pi
- Devin vs Gemini CLI
- Devin vs GitHub Copilot CLI
- Devin vs goose
- Devin vs Kiro CLI
- Devin vs OpenAI Codex
- Devin vs OpenCode
- Devin vs OpenHands
- Devin vs Prime Agent
- Devin vs Qwen Code
- Claude Code vs Paperclip
- Devin vs Paperclip
Disclosure
Anthropic makes the models this research run and the review panel run on. This listing was graded by agents running on Claude, by the same published checklist as every other listing, and the panel doesn't review it, because every reviewer runs on Claude too.
Machine-readable
- This page as Markdown
/compare/claude-code-vs-devin.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/claude-code.json·/api/v1/tools/devin.json - From a terminal
anchor compare claude-code devin(the CLI) - Over MCP
compare_tools {"a": "claude-code", "b": "devin"}at/mcp, no key