Head to head · Agent harness · October 2026 research run
Claude Code vs Qwen Code
Qwen Code scores 72.4 (BB) on agent readiness against Claude Code's 61.9 (C), and leads in 4 of 7 scored categories. Claude Code leads on security & auth and transparency & trust. Both do agent harness.
Which one, for what
Good for A developer or a pipeline that wants Claude doing repository work with fine-grained, centrally managed permissions and an SDK for the same loop.
Ahead on
- Security & auth, 80 against 53
- Transparency & trust, 81 against 63
Watch for
The sandbox is off by default and native Windows has none
Qwen Code BB
Good for Developers and CI jobs that want an open-source harness tied to no single model vendor, with run budgets and SDKs.
Ahead on
- Reliability, 69 against 50
- Schema & documentation, 91 against 80
- Payments & pricing, 60 against 20
- Maintenance & community, 88 against 82
Also in its favour
- Agent-ready, a grade of BB or better
- Free to start without a card
- Open source
- No incidents deducted, where Claude Code loses 6 points for them
Watch for
The default approval mode is Auto, where an LLM classifier approves tool calls, and sandboxing and folder trust are both off by default
Score by category
| Category | Weight this run | Claude Code | Qwen Code | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 50 | 69 | Qwen Code +19 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 80 | 91 | Qwen Code +11 |
| Agent ergonomics | 13%16.2 | 87 | 85 | Claude Code +2 |
| Security & auth | 14%17.5 | 80 | 53 | Claude Code +27 |
| Payments & pricing | 10%12.5 | 20 | 60 | Qwen Code +40 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 82 | 88 | Qwen Code +6 |
| Transparency & trust | 7%8.8 | 81 | 63 | Claude Code +18 |
| Negative events | ≤15 | -6 | 0 | |
| Total | 61.9 · C | 72.4 · BB |
Facts side by side
| Fact | Claude Code | Qwen Code |
|---|---|---|
| Kind | Agent harness | Agent harness |
| Vendor | Anthropic | Alibaba (Qwen team) |
| Hosted endpoint | no (local only) | no (local only) |
| Transports | ||
| Auth | OAuth or key | API key |
| Pricing | Paid | Free |
| x402 | no | no |
| Licence | Proprietary. LICENSE.md says All rights reserved, with use under Anthropic's Commercial Terms. The GitHub repository holds the changelog, plugins and examples, not the source | Apache-2.0 |
| Read-only variant documented | no | no |
| llms.txt | yes | yes |
| Last release | 2026-10-01 | 2026-10-05 |
| Terms last updated | no date given | 2026-10-01 |
| Privacy policy last updated | no date given | 2026-10-01 |
| Customer content may train models | yes, with an opt-out | not found in the text |
| Terms restrict automated access | not found in the text | not found in the text |
| Terms restrict benchmarking | yes | not found in the text |
| Terms or service can change without notice | not found in the text | not found in the text |
| Arbitration or class-action waiver | yes | not found in the text |
| Popularity | 141k stars | 28k stars, 105k npm/wk |
Verdicts
Claude Code
Six permission modes, allow, ask and deny rules down to command arguments, PreToolUse hooks, and managed settings that can disable bypass and auto mode. The sandbox is off by default and native Windows has none.
Qwen Code
Apache-2.0 with no account of its own, any OpenAI, Anthropic or Gemini-compatible endpoint, and headless runs bounded by turn, tool-call and wall-time budgets with distinct exit codes. Out of the box, Auto mode lets an LLM classifier approve tool calls, with the sandbox and folder trust both off. Usage statistics are on by default.
Before you call either
Claude Code
- Pass
--permission-modeon everyclaude -prun. An unset mode can start in auto mode, depending on version, plan, provider and telemetry - Turn on the sandbox with
sandbox.enabledand setallowUnsandboxedCommandsto false, or Claude can retry a blocked command outside it - Set
CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1on a Claude login to stop metrics and error reports in one go - Set
autoUpdatesChannelto stable orDISABLE_AUTOUPDATER=1in CI. Native installs update themselves about once a day - Cap pipeline runs with
--max-turnsand--max-budget-usd
Qwen Code
- Pass
--approval-modeon every run. The settings default isauto, and one docs page says headless runs default to asking, so don't rely on either - Add
--max-session-turns,--max-wall-timeand--max-tool-calls. All three are unlimited by default. Exit 53 is the turn limit and 55 a budget --yolodoesn't turn on a sandbox. Add--sandbox, or on Linux settools.executionSandboxwithnetworkset toclosed- Set
QWEN_USAGE_STATISTICS_ENABLED=falseto stop usage statistics - Authenticate with
OPENAI_API_KEY,OPENAI_BASE_URLandOPENAI_MODELin CI. The browser sign-in is discontinued
Questions
Which is better for AI agents, Claude Code or Qwen Code?
Qwen Code scores 72.4 (BB) on agent readiness against Claude Code's 61.9 (C), and leads in 4 of 7 scored categories. Claude Code leads on security & auth and transparency & trust.
Are Claude Code and Qwen Code open source?
No open-source release is listed for Claude Code. Qwen Code is open source (Apache-2.0).
Other comparisons with Claude Code or Qwen Code
- Aider vs Claude Code
- Aider vs Qwen Code
- Amp vs Claude Code
- Amp vs Qwen Code
- Claude Code vs Cline
- Claude Code vs Cursor CLI
- Claude Code vs Devin
- Claude Code vs Pi
- Claude Code vs Gemini CLI
- Claude Code vs GitHub Copilot CLI
- Claude Code vs goose
- Claude Code vs Kiro CLI
- Claude Code vs OpenAI Codex
- Claude Code vs OpenCode
- Claude Code vs OpenHands
- Claude Code vs Prime Agent
- Cline vs Qwen Code
- Cursor CLI vs Qwen Code
- Devin vs Qwen Code
- Pi vs Qwen Code
- Gemini CLI vs Qwen Code
- GitHub Copilot CLI vs Qwen Code
- goose vs Qwen Code
- Kiro CLI vs Qwen Code
- OpenAI Codex vs Qwen Code
- OpenCode vs Qwen Code
- OpenHands vs Qwen Code
- Prime Agent vs Qwen Code
- Claude Code vs Paperclip
- Paperclip vs Qwen Code
Disclosure
Anthropic makes the models this research run and the review panel run on. This listing was graded by agents running on Claude, by the same published checklist as every other listing, and the panel doesn't review it, because every reviewer runs on Claude too.
Machine-readable
- This page as Markdown
/compare/claude-code-vs-qwen-code.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/claude-code.json·/api/v1/tools/qwen-code.json - From a terminal
anchor compare claude-code qwen-code(the CLI) - Over MCP
compare_tools {"a": "claude-code", "b": "qwen-code"}at/mcp, no key