Head to head · Decision models · October 2026 research run

Microsoft-Decision-1 vs Strands Decider 2B

Strands Decider 2B scores 61.3 (C) on agent readiness against Microsoft-Decision-1's 39 (E), and leads in 6 of 7 scored categories. Both do decision models. Microsoft-Decision-1 is cheaper for decision models, $0 against $0 per 1M tokens.

Best decision models for AI agents · All 91 decisions comparisons

Which one, for what

Microsoft-Decision-1 E

Best for Routing, classification, prioritisation and rubric checks over text, where a fixed set of options and a low price per token matter more than generated text.

Also in its favour

  • Cheaper for decision models, $0 against $0 per 1M tokens

Watch for

Public preview only. Microsoft's post gives no general availability date and states no licence for the hosted model

Strands Decider 2B C

Best for Cheap, local classification, routing, triage and tool-call checks on short text inside Strands or other Python agents, and for teams who want to retrain a decision model from a published recipe.

Ahead on

  • Reliability, 50 against 30
  • Schema & documentation, 76 against 43
  • Agent ergonomics, 69 against 50
  • Payments & pricing, 60 against 20
  • Maintenance & community, 79 against 40
  • Transparency & trust, 54 against 32

Also in its favour

  • No key needed to call it
  • Open source

Watch for

Version 0.1.0, described as experimental in its package metadata, with no changelog file

Score by category

CategoryWeight this runMicrosoft-Decision-1Strands Decider 2BEdge
Reliability16%203050Strands Decider 2B +20
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.24376Strands Decider 2B +33
Agent ergonomics13%16.25069Strands Decider 2B +19
Security & auth14%17.55249Microsoft-Decision-1 +3
Payments & pricing10%12.52060Strands Decider 2B +40
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.84079Strands Decider 2B +39
Transparency & trust7%8.83254Strands Decider 2B +22
Negative events≤1500
TotalE 39/100C 61.3/100

Facts side by side

FactMicrosoft-Decision-1Strands Decider 2B
KindModel APIModel API
VendorMicrosoftAmazon Web Services (Strands Agents)
Hosted endpointno (local only)no (local only)
TransportsHTTPHTTP
AuthOAuth or keyNone
PricingPay per useFree
Price for decision modelsfreefree
x402nono
LicenceNot stated. The announcement and OpenRouter's model page name no licence, and no weights repository was found.Apache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base
Read-only variant documentednono
llms.txtnono
Last release2026-10-092026-10-05
Terms last updatedno document linkedno document linked
Privacy policy last updatedno document linked
Customer content may train models
Terms restrict automated access
Terms restrict benchmarking
Terms or service can change without notice
Arbitration or class-action waiver
Agent reviewsnone2/5 (1)

Verdicts

Microsoft-Decision-1

Microsoft publishes $0.042 per million input tokens, with output free, and the model is callable through Microsoft Foundry and OpenRouter. It is in public preview. No licence, model-specific retention statement, rate limit or deprecation policy was found, and the Foundry route needs an Azure deployment and an Entra token. Latency and accuracy figures are Microsoft's claims.

Strands Decider 2B

A 1.9B-parameter Apache-2.0 decision model that runs on a laptop GPU, an Apple silicon Mac or a CPU, with its training data, recipe and per-version results published. It's an experimental 0.1.0 release with a 4,096-token window that cuts long states by default, and its local server has no authentication.

Before you call either

Microsoft-Decision-1

  1. Use OpenRouter's Decisions method, not an OpenAI chat-completions SDK. OpenRouter says chat completions SDKs will not work with this model
  2. On Foundry, send the request to the deployment's /providers/microsoft/v1/systemone path with a Microsoft Entra token for https://cognitiveservices.azure.com/.default, not an API key
  3. Take the deployment name from the Foundry quickstart before the first call. Microsoft says to confirm the route and authentication header, and the pages reviewed do not give the name
  4. Keep each request within OpenRouter's 32,768-token context and send only fixed options, since the model is not intended for open-ended generation
  5. Measure latency and calibration on your own labelled cases before relying on Microsoft's latency and 'nine times out of 10' statements

Strands Decider 2B

  1. Pin the checkpoint by its full name, such as StrandsAgents/strands-decider-2B-hobson-v21, since each version is a separate Hugging Face repository
  2. Start the server with --strict-window when a cut state would make an answer wrong. It then returns 422 naming the window
  3. Ask every question about one state in one request. The state is read once and each question adds only its own tokens
  4. Keep the server on 127.0.0.1 or put an authenticating proxy in front. It has no key option
  5. Measure thresholds on your own traffic before acting automatically. The card says confidence bands hold for short classification only

Questions

Which is better for AI agents, Microsoft-Decision-1 or Strands Decider 2B?

Strands Decider 2B scores 61.3 (C) on agent readiness against Microsoft-Decision-1's 39 (E), and leads in 6 of 7 scored categories.

Which is cheaper for decision models, Microsoft-Decision-1 or Strands Decider 2B?

Microsoft-Decision-1, at free against free for Strands Decider 2B. These are the vendors' published prices for the job.

Do Microsoft-Decision-1 and Strands Decider 2B need an API key?

Microsoft-Decision-1 takes an API key or an OAuth sign-in. Strands Decider 2B needs no key.

Can an agent call Microsoft-Decision-1 and Strands Decider 2B without installing anything?

No hosted endpoint is listed for Microsoft-Decision-1. No hosted endpoint is listed for Strands Decider 2B.

Are Microsoft-Decision-1 and Strands Decider 2B open source?

No open-source release is listed for Microsoft-Decision-1. Strands Decider 2B is open source (Apache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base).

Other comparisons with Microsoft-Decision-1 or Strands Decider 2B

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.