Head to head · Agent harnesses · October 2026 research run
Claude Code vs Droid CLI
Claude Code scores 61.9 (C) on agent readiness against Droid CLI's 59.1 (C), and leads in every scored category. Both do agent harnesses.
Best agent harnesses and coding agents · All 167 harnesses comparisons
Which one, for what
Good for A developer or a pipeline that wants Claude doing repository work with fine-grained, centrally managed permissions and an SDK for the same loop.
Ahead on
- Agent ergonomics, 87 against 72
- Security & auth, 80 against 63
- Payments & pricing, 20 against 10
- Transparency & trust, 81 against 64
Watch for
The sandbox is off by default and native Windows has none
Good for Teams that want a terminal coding agent with tiered autonomy, command rules and an optional sandbox, and that will pay for a Factory plan.
Also in its favour
- No incidents deducted, where Claude Code loses 6 points for them
Watch for
No free tier or trial was found. Plans start at $20 a month and included usage is stated only as rolling rate limits without numbers
Score by category
| Category | Weight this run | Claude Code | Droid CLI | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 50 | 49 | Claude Code +1 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 80 | 79 | Claude Code +1 |
| Agent ergonomics | 13%16.2 | 87 | 72 | Claude Code +15 |
| Security & auth | 14%17.5 | 80 | 63 | Claude Code +17 |
| Payments & pricing | 10%12.5 | 20 | 10 | Claude Code +10 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 82 | 79 | Claude Code +3 |
| Transparency & trust | 7%8.8 | 81 | 64 | Claude Code +17 |
| Negative events | ≤15 | -6 | 0 | |
| Total | 61.9 · C | 59.1 · C |
Facts side by side
| Fact | Claude Code | Droid CLI |
|---|---|---|
| Kind | Agent harness | Agent harness |
| Vendor | Anthropic | Factory |
| Hosted endpoint | no (local only) | no (local only) |
| Transports | ||
| Auth | OAuth or key | OAuth or key |
| Pricing | Paid | Paid |
| x402 | no | no |
| Licence | Proprietary. LICENSE.md says All rights reserved, with use under Anthropic's Commercial Terms. The GitHub repository holds the changelog, plugins and examples, not the source | Proprietary, under Factory's Terms and Conditions. The droid npm package is marked UNLICENSED and the public repository holds documentation only. The Python SDK is Apache-2.0 |
| Read-only variant documented | no | yes |
| llms.txt | yes | yes |
| Last release | 2026-10-01 | 2026-10-08 |
| Terms last updated | no date given | 2026-07-14 |
| Privacy policy last updated | no date given | 2026-08-20 |
| Customer content may train models | yes, with an opt-out | not found in the text |
| Terms restrict automated access | not found in the text | not found in the text |
| Terms restrict benchmarking | yes | yes |
| Terms or service can change without notice | not found in the text | not found in the text |
| Arbitration or class-action waiver | yes | yes |
| Popularity | 141k stars | 49 stars, 7.8k npm/wk |
Verdicts
Claude Code
Six permission modes, allow, ask and deny rules down to command arguments, PreToolUse hooks, and managed settings that can disable bypass and auto mode. The sandbox is off by default and native Windows has none.
Droid CLI
droid exec is read-only unless --auto raises the autonomy level, with command rules, an optional kernel-enforced sandbox, JSON output and documented exit codes. No free tier, status page or security.txt was found, the source is closed, and usage metrics go to Factory by default.
Before you call either
Claude Code
- Pass
--permission-modeon everyclaude -prun. An unset mode can start in auto mode, depending on version, plan, provider and telemetry - Turn on the sandbox with
sandbox.enabledand setallowUnsandboxedCommandsto false, or Claude can retry a blocked command outside it - Set
CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1on a Claude login to stop metrics and error reports in one go - Set
autoUpdatesChannelto stable orDISABLE_AUTOUPDATER=1in CI. Native installs update themselves about once a day - Cap pipeline runs with
--max-turnsand--max-budget-usd
Droid CLI
- Set
FACTORY_API_KEY(startsfk-) from the API keys page in Factory settings for headless runs. Interactive use signs in through a browser - Start with
droid execand no flags for analysis. Add--auto lowfor edits and--auto mediumfor installs, tests and local commits - Use
--output-format jsonand readis_error,num_turnsandsession_id. Treat a non-zero exit code as failure - Set
sandbox.enabledto true before running on untrusted code. Indroid execa sandbox violation is denied without a prompt - Pin the version in CI with
npm install -g droid@<version>, or setFACTORY_DROID_AUTO_UPDATE_ENABLED=falseon standalone installs, which update themselves
Questions
Which is better for AI agents, Claude Code or Droid CLI?
Claude Code scores 61.9 (C) on agent readiness against Droid CLI's 59.1 (C), and leads in every scored category.
Other comparisons with Claude Code or Droid CLI
- Aider vs Claude Code
- Aider vs Droid CLI
- Amp vs Claude Code
- Amp vs Droid CLI
- Claude Code vs Cline
- Claude Code vs Cursor CLI
- Claude Code vs Devin
- Claude Code vs Pi
- Claude Code vs Gemini CLI
- Claude Code vs GitHub Copilot CLI
- Claude Code vs goose
- Claude Code vs Kilo Code CLI
- Claude Code vs Kiro CLI
- Claude Code vs OpenAI Codex
- Claude Code vs OpenCode
- Claude Code vs OpenHands
- Claude Code vs Prime Agent
- Claude Code vs Qwen Code
- Cline vs Droid CLI
- Cursor CLI vs Droid CLI
- Devin vs Droid CLI
- Droid CLI vs Pi
- Droid CLI vs Gemini CLI
- Droid CLI vs GitHub Copilot CLI
- Droid CLI vs goose
- Droid CLI vs Kilo Code CLI
- Droid CLI vs Kiro CLI
- Droid CLI vs OpenAI Codex
- Droid CLI vs OpenCode
- Droid CLI vs OpenHands
- Droid CLI vs Prime Agent
- Droid CLI vs Qwen Code
- Claude Code vs Paperclip
- Droid CLI vs Paperclip
Disclosure
Anthropic makes the models this research run and the review panel run on. This listing was graded by agents running on Claude, by the same published checklist as every other listing, and the panel doesn't review it, because every reviewer runs on Claude too.
Machine-readable
- This page as Markdown
/compare/claude-code-vs-droid-cli.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/claude-code.json·/api/v1/tools/droid-cli.json - From a terminal
anchor compare claude-code droid-cli(the CLI) - Over MCP
compare_tools {"a": "claude-code", "b": "droid-cli"}at/mcp, no key