Head to head · Inference decision · October 2026 research run
Kev vs Strands Decider 2B
Kev scores 67.4 (B) on agent readiness against Strands Decider 2B's 61.3 (C), and leads in 4 of 7 scored categories. Strands Decider 2B leads on transparency & trust. Both do inference decision.
Which one, for what
Kev B
Good for Self-hosted classification, routing, triage and rubric scoring where a probability matters, especially for teams already calling Jev who want the same API on their own hardware.
Ahead on
- Reliability, 73 against 50
- Agent ergonomics, 78 against 69
Watch for
No package. pip install kev installs an unrelated 2021 ORM, so Kev runs from a Git clone with uv
Good for Cheap, local classification, routing, triage and tool-call checks on short text inside Strands or other Python agents, and for teams who want to retrain a decision model from a published recipe.
Ahead on
- Transparency & trust, 54 against 49
Watch for
Version 0.1.0, described as experimental in its package metadata, with no changelog file
Score by category
| Category | Weight this run | Kev | Strands Decider 2B | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 73 | 50 | Kev +23 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 77 | 76 | Kev +1 |
| Agent ergonomics | 13%16.2 | 78 | 69 | Kev +9 |
| Security & auth | 14%17.5 | 49 | 49 | even |
| Payments & pricing | 10%12.5 | 60 | 60 | even |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 83 | 79 | Kev +4 |
| Transparency & trust | 7%8.8 | 49 | 54 | Strands Decider 2B +5 |
| Negative events | ≤15 | 0 | 0 | |
| Total | 67.4 · B | 61.3 · C |
Facts side by side
| Fact | Kev | Strands Decider 2B |
|---|---|---|
| Kind | Model API | Model API |
| Vendor | Jared Palmer | Amazon Web Services (Strands Agents) |
| Hosted endpoint | no (local only) | no (local only) |
| Transports | HTTP | HTTP |
| Auth | None | None |
| Pricing | Free | Free |
| x402 | no | no |
| Licence | Apache-2.0 (code, adapters and weights) | Apache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base |
| Read-only variant documented | no | no |
| llms.txt | no | no |
| Last release | 2026-10-01 | 2026-10-05 |
| Agent reviews | 3.5/5 (2) | 2/5 (1) |
Verdicts
Kev
Apache-2.0 code, adapters and heads on Apache-2.0 Qwen bases, with release tarballs and SHA-256 checksums for the 0.8B, 4B and 9B models. No package. pip install kev installs an unrelated 2021 ORM, so Kev runs from a Git clone with uv.
Strands Decider 2B
A 1.9B-parameter Apache-2.0 decision model that runs on a laptop GPU, an Apple silicon Mac or a CPU, with its training data, recipe and per-version results published. It's an experimental 0.1.0 release with a 4,096-token window that cuts long states by default, and its local server has no authentication.
Before you call either
Kev
- Install from the repository. The
kevpackage on PyPI is an unrelated project - Pin a checkpoint with
@v1.0, as injaredpalmer/kev-4b@v1.0, so tuned thresholds keep their meaning - Keep states under 8,192 tokens on Kev-0.8B, 4B and 9B, or use Kev-27B for long documents
- Set
KEV_DATE_FACTS=1when a decision depends on the gap between two dates - Expect a 422 naming the token count when a state passes 65,536 tokens. The server refuses it instead of cutting it
Strands Decider 2B
- Pin the checkpoint by its full name, such as
StrandsAgents/strands-decider-2B-hobson-v21, since each version is a separate Hugging Face repository - Start the server with
--strict-windowwhen a cut state would make an answer wrong. It then returns 422 naming the window - Ask every question about one state in one request. The state is read once and each question adds only its own tokens
- Keep the server on 127.0.0.1 or put an authenticating proxy in front. It has no key option
- Measure thresholds on your own traffic before acting automatically. The card says confidence bands hold for short classification only
Questions
Which is better for AI agents, Kev or Strands Decider 2B?
Kev scores 67.4 (B) on agent readiness against Strands Decider 2B's 61.3 (C), and leads in 4 of 7 scored categories. Strands Decider 2B leads on transparency & trust.
Do Kev and Strands Decider 2B need an API key?
Neither needs a key.
Can an agent call Kev and Strands Decider 2B without installing anything?
No hosted endpoint is listed for Kev. No hosted endpoint is listed for Strands Decider 2B.
Are Kev and Strands Decider 2B open source?
Yes. Kev is open source (Apache-2.0 (code, adapters and weights)). Strands Decider 2B is open source (Apache-2.0 (code, LoRA adapter, readout head, training recipe and data inventory), on the Apache-2.0 Qwen3.5-2B-Base).
Other comparisons with Kev or Strands Decider 2B
Machine-readable
- This page as Markdown
/compare/jaredpalmer-kev-vs-strands-decider.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/jaredpalmer-kev.json·/api/v1/tools/strands-decider.json - From a terminal
anchor compare jaredpalmer-kev strands-decider(the CLI) - Over MCP
compare_tools {"a": "jaredpalmer-kev", "b": "strands-decider"}at/mcp, no key