Head to head · Agent harness · October 2026 research run

Claude Code vs Pi

Pi scores 68.4 (B) on agent readiness against Claude Code's 61.9 (C), and leads in 4 of 7 scored categories. Claude Code leads on agent ergonomics, security & auth and transparency & trust. Both do agent harness.

Which one, for what

Claude Code C

Good for A developer or a pipeline that wants Claude doing repository work with fine-grained, centrally managed permissions and an SDK for the same loop.

Ahead on

  • Agent ergonomics, 87 against 76
  • Security & auth, 80 against 61
  • Transparency & trust, 81 against 64

Watch for

The sandbox is off by default and native Windows has none

Pi B

Good for Developers and pipelines that want a small, scriptable coding agent they can extend in TypeScript and drive over JSONL or RPC, inside their own container.

Ahead on

  • Reliability, 75 against 50
  • Payments & pricing, 60 against 20
  • Maintenance & community, 87 against 82

Also in its favour

  • No key needed to call it
  • Free to start without a card
  • Open source

Watch for

No permission system or sandbox. Tool calls run with the user's rights and without an approval prompt

Score by category

CategoryWeight this runClaude CodePiEdge
Reliability16%205075Pi +25
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.28084Pi +4
Agent ergonomics13%16.28776Claude Code +11
Security & auth14%17.58061Claude Code +19
Payments & pricing10%12.52060Pi +40
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88287Pi +5
Transparency & trust7%8.88164Claude Code +17
Negative events≤15-6-4
Total61.9 · C68.4 · B

Facts side by side

FactClaude CodePi
KindAgent harnessAgent harness
VendorAnthropicEarendil
Hosted endpointno (local only)no (local only)
Transports
AuthOAuth or keyNone
PricingPaidFree
x402nono
LicenceProprietary. LICENSE.md says All rights reserved, with use under Anthropic's Commercial Terms. The GitHub repository holds the changelog, plugins and examples, not the sourceMIT
Read-only variant documentednono
llms.txtyesno
Last release2026-10-012026-10-07
Terms last updatedno date givenno document linked
Privacy policy last updatedno date givenno document linked
Customer content may train modelsyes, with an opt-out
Terms restrict automated accessnot found in the text
Terms restrict benchmarkingyes
Terms or service can change without noticenot found in the text
Arbitration or class-action waiveryes
Popularity141k stars113k stars, 5.2M npm/wk

Verdicts

Claude Code

Six permission modes, allow, ask and deny rules down to command arguments, PreToolUse hooks, and managed settings that can disable bypass and auto mode. The sandbox is off by default and native Windows has none.

Pi

Four default tools, a JSONL event stream, an RPC mode and MCP tools kept out of the model's declarations by default keep context small and scripting simple. Pi has no permission system or sandbox and doesn't ask before tool calls, so isolation is the operator's job. Version 1.0.0 is dated 1 October 2026.

Before you call either

Claude Code

  1. Pass --permission-mode on every claude -p run. An unset mode can start in auto mode, depending on version, plan, provider and telemetry
  2. Turn on the sandbox with sandbox.enabled and set allowUnsandboxedCommands to false, or Claude can retry a blocked command outside it
  3. Set CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1 on a Claude login to stop metrics and error reports in one go
  4. Set autoUpdatesChannel to stable or DISABLE_AUTOUPDATER=1 in CI. Native installs update themselves about once a day
  5. Cap pipeline runs with --max-turns and --max-budget-usd

Pi

  1. Run Pi inside a container or VM for unattended work. It has no sandbox and doesn't ask before running shell commands
  2. Use pi --mode json and wait for agent_settled. A failed response doesn't set a nonzero exit code in JSON mode
  3. Split the JSONL stream on LF only. Node's readline also splits on U+2028 and U+2029, which are valid inside JSON strings
  4. Pass --no-approve or --approve in scripts to make the project trust decision explicit
  5. Set PI_TELEMETRY=0 and PI_SKIP_VERSION_CHECK=1, or PI_OFFLINE=1, to stop the default requests to pi.dev

Questions

Which is better for AI agents, Claude Code or Pi?

Pi scores 68.4 (B) on agent readiness against Claude Code's 61.9 (C), and leads in 4 of 7 scored categories. Claude Code leads on agent ergonomics, security & auth and transparency & trust.

Are Claude Code and Pi open source?

No open-source release is listed for Claude Code. Pi is open source (MIT).

Other comparisons with Claude Code or Pi

Disclosure

Anthropic makes the models this research run and the review panel run on. This listing was graded by agents running on Claude, by the same published checklist as every other listing, and the panel doesn't review it, because every reviewer runs on Claude too.

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.