Head to head · Agent harness · October 2026 research run
Devin vs Qwen Code
Qwen Code scores 72.4 (BB) on agent readiness against Devin's 55.7 (C), and leads in 5 of 7 scored categories. Devin leads on security & auth and transparency & trust. Both do agent harness.
Which one, for what
Devin C
Good for A team that wants to hand whole tasks to a cloud agent and collect pull requests, driven from a pipeline or another agent.
Ahead on
- Security & auth, 72 against 53
- Transparency & trust, 69 against 63
Also in its favour
- A hosted endpoint, with nothing to install
Watch for
www.devinstatus.com lists three critical incidents (22 July, 13 August, 24 September 2026) and six major ones on the cloud agent or web app since 10 July 2026
Qwen Code BB
Good for Developers and CI jobs that want an open-source harness tied to no single model vendor, with run budgets and SDKs.
Ahead on
- Reliability, 69 against 30
- Schema & documentation, 91 against 80
- Agent ergonomics, 85 against 62
- Payments & pricing, 60 against 20
- Maintenance & community, 88 against 63
Also in its favour
- Agent-ready, a grade of BB or better
- Free to start without a card
- Open source
Watch for
The default approval mode is Auto, where an LLM classifier approves tool calls, and sandboxing and folder trust are both off by default
Score by category
| Category | Weight this run | Devin | Qwen Code | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 30 | 69 | Qwen Code +39 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 80 | 91 | Qwen Code +11 |
| Agent ergonomics | 13%16.2 | 62 | 85 | Qwen Code +23 |
| Security & auth | 14%17.5 | 72 | 53 | Devin +19 |
| Payments & pricing | 10%12.5 | 20 | 60 | Qwen Code +40 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 63 | 88 | Qwen Code +25 |
| Transparency & trust | 7%8.8 | 69 | 63 | Devin +6 |
| Negative events | ≤15 | 0 | 0 | |
| Total | 55.7 · C | 72.4 · BB |
Facts side by side
| Fact | Devin | Qwen Code |
|---|---|---|
| Kind | Agent harness | Agent harness |
| Vendor | Cognition AI, Inc. | Alibaba (Qwen team) |
| Hosted endpoint | https://mcp.devin.ai/mcp | no (local only) |
| Transports | HTTP | |
| Auth | API key | API key |
| Pricing | Freemium | Free |
| x402 | no | no |
| Licence | Proprietary service under Cognition's Platform Terms of Service | Apache-2.0 |
| Tools exposed | 13 | none |
| Read-only variant documented | yes | no |
| llms.txt | yes | yes |
| Last release | 2026-10-07 | 2026-10-05 |
| Terms last updated | 2026-06-30 | 2026-10-01 |
| Privacy policy last updated | 2026-03-09 | 2026-10-01 |
| Customer content may train models | yes, with an opt-out | not found in the text |
| Terms restrict automated access | not found in the text | not found in the text |
| Terms restrict benchmarking | yes | not found in the text |
| Terms or service can change without notice | not found in the text | not found in the text |
| Arbitration or class-action waiver | yes | not found in the text |
| Popularity | none | 28k stars, 105k npm/wk |
Verdicts
Devin
The v3 API is well specified, with a public OpenAPI 3.1 file covering 239 operations, problem+json errors, cursor pagination and service-user keys tied to roles. The status page records three critical and several major incidents on the cloud agent between 22 July and 24 September 2026, and no rate limit figures or retry guidance were found in the reviewed documentation.
Qwen Code
Apache-2.0 with no account of its own, any OpenAI, Anthropic or Gemini-compatible endpoint, and headless runs bounded by turn, tool-call and wall-time budgets with distinct exit codes. Out of the box, Auto mode lets an LLM classifier approve tool calls, with the sandbox and folder trust both off. Usage statistics are on by default.
Before you call either
Devin
- Use a
cog_service-user key with the Member role. Legacyapk_keys fail against v3 and the MCP server with 401 or 403 - Set
max_acu_limiton every session you create. Usage is metered by the work done and has no published unit rate - Session creation is not idempotent in v3. After a timeout, list sessions by tag before creating again
- Enterprise keys and personal access tokens must send
X-Org-Idto the MCP server. Organisation-scoped keys resolve it automatically - Create scheduled work as an automation. POST to the schedules endpoint returns 403 for migrated organisations since 24 September 2026
Qwen Code
- Pass
--approval-modeon every run. The settings default isauto, and one docs page says headless runs default to asking, so don't rely on either - Add
--max-session-turns,--max-wall-timeand--max-tool-calls. All three are unlimited by default. Exit 53 is the turn limit and 55 a budget --yolodoesn't turn on a sandbox. Add--sandbox, or on Linux settools.executionSandboxwithnetworkset toclosed- Set
QWEN_USAGE_STATISTICS_ENABLED=falseto stop usage statistics - Authenticate with
OPENAI_API_KEY,OPENAI_BASE_URLandOPENAI_MODELin CI. The browser sign-in is discontinued
Questions
Which is better for AI agents, Devin or Qwen Code?
Qwen Code scores 72.4 (BB) on agent readiness against Devin's 55.7 (C), and leads in 5 of 7 scored categories. Devin leads on security & auth and transparency & trust.
Are Devin and Qwen Code open source?
No open-source release is listed for Devin. Qwen Code is open source (Apache-2.0).
Other comparisons with Devin or Qwen Code
- Aider vs Devin
- Aider vs Qwen Code
- Amp vs Devin
- Amp vs Qwen Code
- Claude Code vs Devin
- Claude Code vs Qwen Code
- Cline vs Devin
- Cline vs Qwen Code
- Cursor CLI vs Devin
- Cursor CLI vs Qwen Code
- Devin vs Pi
- Devin vs Gemini CLI
- Devin vs GitHub Copilot CLI
- Devin vs goose
- Devin vs Kiro CLI
- Devin vs OpenAI Codex
- Devin vs OpenCode
- Devin vs OpenHands
- Devin vs Prime Agent
- Pi vs Qwen Code
- Gemini CLI vs Qwen Code
- GitHub Copilot CLI vs Qwen Code
- goose vs Qwen Code
- Kiro CLI vs Qwen Code
- OpenAI Codex vs Qwen Code
- OpenCode vs Qwen Code
- OpenHands vs Qwen Code
- Prime Agent vs Qwen Code
- Devin vs Paperclip
- Paperclip vs Qwen Code
Machine-readable
- This page as Markdown
/compare/devin-vs-qwen-code.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/devin.json·/api/v1/tools/qwen-code.json - From a terminal
anchor compare devin qwen-code(the CLI) - Over MCP
compare_tools {"a": "devin", "b": "qwen-code"}at/mcp, no key