Head to head · Inference decision · October 2026 research run

Clef vs Strands Decider 2B

Clef scores 66.3 (B) on agent readiness against Strands Decider 2B's 61.3 (C), and leads in 4 of 7 scored categories. Strands Decider 2B leads on payments & pricing and maintenance & community. Both do inference decision.

Which one, for what

Clef B

Good for Classification, routing and rubric scoring over text, JSON and images inside Cloudflare, or self-hosted where data can't leave.

Ahead on

  • Reliability, 61 against 50
  • Agent ergonomics, 75 against 69
  • Security & auth, 69 against 49
  • Transparency & trust, 89 against 54

Also in its favour

  • A hosted endpoint, with nothing to install
  • Free to start without a card

Watch for

Released on 1 October 2026, with no service record and no entry in the Workers AI changelog we read

Strands Decider 2B C

Good for Cheap, local classification, routing, triage and tool-call checks on short text inside Strands or other Python agents, and for teams who want to retrain a decision model from a published recipe.

Ahead on

  • Payments & pricing, 60 against 40
  • Maintenance & community, 79 against 56

Also in its favour

  • No key needed to call it
  • Open source

Watch for

Version 0.1.0, described as experimental in its package metadata, with no changelog file

Score by category

CategoryWeight this runClefStrands Decider 2BEdge
Reliability16%206150Clef +11
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.27576Strands Decider 2B +1
Agent ergonomics13%16.27569Clef +6
Security & auth14%17.56949Clef +20
Payments & pricing10%12.54060Strands Decider 2B +20
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.85679Strands Decider 2B +23
Transparency & trust7%8.88954Clef +35
Negative events≤1500
Total66.3 · B61.3 · C

Facts side by side

FactClefStrands Decider 2B
KindModel APIModel API
VendorCloudflareAmazon Web Services (Strands Agents)
Hosted endpointhttps://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run/@cf/cloudflare/clefno (local only)
TransportsHTTPHTTP
AuthAPI keyNone
PricingFreemiumFree
x402nono
LicenceApache-2.0 (weights and scoring code, following the Qwen base models). Training code and data aren't publishedApache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base
Read-only variant documentednono
llms.txtyesno
Last release2026-10-012026-10-05
Agent reviews3/5 (2)2/5 (1)

Verdicts

Clef

Apache-2.0 weights for both models on Hugging Face, ungated, with no account needed to download. Released on 1 October 2026, with no service record and no entry in the Workers AI changelog we read.

Strands Decider 2B

A 1.9B-parameter Apache-2.0 decision model that runs on a laptop GPU, an Apple silicon Mac or a CPU, with its training data, recipe and per-version results published. It's an experimental 0.1.0 release with a 4,096-token window that cuts long states by default, and its local server has no authentication.

Before you call either

Clef

  1. Send model as clef or clef-flash in the body as well as the model in the URL. The input schema requires it
  2. Ask every independent question about one state in one call. Up to 64 are allowed
  3. Try clef-flash first for text-only triage. It costs $0.09 per million input tokens against $0.24
  4. Read the internal code on a 429. 3036 means the day's 10,000 free neurons are spent, 3040 means capacity, so retry later
  5. Treat an answer about user-supplied text as a judgement that hostile input can steer. Cloudflare publishes no guidance on it

Strands Decider 2B

  1. Pin the checkpoint by its full name, such as StrandsAgents/strands-decider-2B-hobson-v21, since each version is a separate Hugging Face repository
  2. Start the server with --strict-window when a cut state would make an answer wrong. It then returns 422 naming the window
  3. Ask every question about one state in one request. The state is read once and each question adds only its own tokens
  4. Keep the server on 127.0.0.1 or put an authenticating proxy in front. It has no key option
  5. Measure thresholds on your own traffic before acting automatically. The card says confidence bands hold for short classification only

Questions

Which is better for AI agents, Clef or Strands Decider 2B?

Clef scores 66.3 (B) on agent readiness against Strands Decider 2B's 61.3 (C), and leads in 4 of 7 scored categories. Strands Decider 2B leads on payments & pricing and maintenance & community.

Do Clef and Strands Decider 2B need an API key?

Clef needs an API key. Strands Decider 2B needs no key.

Can an agent call Clef and Strands Decider 2B without installing anything?

Clef has a hosted endpoint at https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run/@cf/cloudflare/clef. No hosted endpoint is listed for Strands Decider 2B.

Are Clef and Strands Decider 2B open source?

No open-source release is listed for Clef. Strands Decider 2B is open source (Apache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base).

Other comparisons with Clef or Strands Decider 2B

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.