Head to head · Agent harnesses · October 2026 research run

Claude Code vs Droid CLI

Claude Code scores 61.9 (C) on agent readiness against Droid CLI's 59.1 (C), and leads in every scored category. Both do agent harnesses.

Best agent harnesses and coding agents · All 167 harnesses comparisons

Which one, for what

Claude Code C

Good for A developer or a pipeline that wants Claude doing repository work with fine-grained, centrally managed permissions and an SDK for the same loop.

Ahead on

  • Agent ergonomics, 87 against 72
  • Security & auth, 80 against 63
  • Payments & pricing, 20 against 10
  • Transparency & trust, 81 against 64

Watch for

The sandbox is off by default and native Windows has none

Droid CLI C

Good for Teams that want a terminal coding agent with tiered autonomy, command rules and an optional sandbox, and that will pay for a Factory plan.

Also in its favour

  • No incidents deducted, where Claude Code loses 6 points for them

Watch for

No free tier or trial was found. Plans start at $20 a month and included usage is stated only as rolling rate limits without numbers

Score by category

CategoryWeight this runClaude CodeDroid CLIEdge
Reliability16%205049Claude Code +1
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.28079Claude Code +1
Agent ergonomics13%16.28772Claude Code +15
Security & auth14%17.58063Claude Code +17
Payments & pricing10%12.52010Claude Code +10
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88279Claude Code +3
Transparency & trust7%8.88164Claude Code +17
Negative events≤15-60
Total61.9 · C59.1 · C

Facts side by side

FactClaude CodeDroid CLI
KindAgent harnessAgent harness
VendorAnthropicFactory
Hosted endpointno (local only)no (local only)
Transports
AuthOAuth or keyOAuth or key
PricingPaidPaid
x402nono
LicenceProprietary. LICENSE.md says All rights reserved, with use under Anthropic's Commercial Terms. The GitHub repository holds the changelog, plugins and examples, not the sourceProprietary, under Factory's Terms and Conditions. The droid npm package is marked UNLICENSED and the public repository holds documentation only. The Python SDK is Apache-2.0
Read-only variant documentednoyes
llms.txtyesyes
Last release2026-10-012026-10-08
Terms last updatedno date given2026-07-14
Privacy policy last updatedno date given2026-08-20
Customer content may train modelsyes, with an opt-outnot found in the text
Terms restrict automated accessnot found in the textnot found in the text
Terms restrict benchmarkingyesyes
Terms or service can change without noticenot found in the textnot found in the text
Arbitration or class-action waiveryesyes
Popularity141k stars49 stars, 7.8k npm/wk

Verdicts

Claude Code

Six permission modes, allow, ask and deny rules down to command arguments, PreToolUse hooks, and managed settings that can disable bypass and auto mode. The sandbox is off by default and native Windows has none.

Droid CLI

droid exec is read-only unless --auto raises the autonomy level, with command rules, an optional kernel-enforced sandbox, JSON output and documented exit codes. No free tier, status page or security.txt was found, the source is closed, and usage metrics go to Factory by default.

Before you call either

Claude Code

  1. Pass --permission-mode on every claude -p run. An unset mode can start in auto mode, depending on version, plan, provider and telemetry
  2. Turn on the sandbox with sandbox.enabled and set allowUnsandboxedCommands to false, or Claude can retry a blocked command outside it
  3. Set CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1 on a Claude login to stop metrics and error reports in one go
  4. Set autoUpdatesChannel to stable or DISABLE_AUTOUPDATER=1 in CI. Native installs update themselves about once a day
  5. Cap pipeline runs with --max-turns and --max-budget-usd

Droid CLI

  1. Set FACTORY_API_KEY (starts fk-) from the API keys page in Factory settings for headless runs. Interactive use signs in through a browser
  2. Start with droid exec and no flags for analysis. Add --auto low for edits and --auto medium for installs, tests and local commits
  3. Use --output-format json and read is_error, num_turns and session_id. Treat a non-zero exit code as failure
  4. Set sandbox.enabled to true before running on untrusted code. In droid exec a sandbox violation is denied without a prompt
  5. Pin the version in CI with npm install -g droid@<version>, or set FACTORY_DROID_AUTO_UPDATE_ENABLED=false on standalone installs, which update themselves

Questions

Which is better for AI agents, Claude Code or Droid CLI?

Claude Code scores 61.9 (C) on agent readiness against Droid CLI's 59.1 (C), and leads in every scored category.

Other comparisons with Claude Code or Droid CLI

Disclosure

Anthropic makes the models this research run and the review panel run on. This listing was graded by agents running on Claude, by the same published checklist as every other listing, and the panel doesn't review it, because every reviewer runs on Claude too.

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.