Head to head · Inference decision · October 2026 research run

Decider vs GLiClass

Decider scores 69.5 (B) on agent readiness against GLiClass's 49.9 (D), and leads in 4 of 7 scored categories. GLiClass leads on transparency & trust. Both do inference decision.

Which one, for what

Decider B

Good for Local classification, routing, triage and checks where a team wants open weights in several sizes and a Jev-shaped route.

Ahead on

  • Reliability, 90 against 43
  • Schema & documentation, 78 against 49
  • Agent ergonomics, 80 against 60
  • Maintenance & community, 87 against 44

Watch for

The local server has no authentication option. It binds to 127.0.0.1 since 1.7.1

GLiClass D

Good for Topic, intent and sentiment routing over a known label set on the owner's own CPU or GPU, where many labels must be scored at once.

Ahead on

  • Transparency & trust, 64 against 46

Watch for

No calibration evidence. The cards report F1 only, and the docs tell users to calibrate thresholds on their own traffic

Score by category

CategoryWeight this runDeciderGLiClassEdge
Reliability16%209043Decider +47
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.27849Decider +29
Agent ergonomics13%16.28060Decider +20
Security & auth14%17.53838even
Payments & pricing10%12.56060even
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88744Decider +43
Transparency & trust7%8.84664GLiClass +18
Negative events≤1500
Total69.5 · B49.9 · D

Facts side by side

FactDeciderGLiClass
KindModel APIModel API
VendorMark Marosi (Mapika)Knowledgator
Hosted endpointno (local only)no (local only)
TransportsHTTPHTTP
AuthNoneNone
PricingFreeFree
x402nono
LicenceApache-2.0 (code and weights)Apache-2.0 (library and the model weights we checked)
Read-only variant documentednono
llms.txtnono
Last release2026-10-072026-07-21
Terms last updatedno document linkedno document linked
Privacy policy last updatedno document linkedno document linked
Customer content may train models
Terms restrict automated access
Terms restrict benchmarking
Terms or service can change without notice
Arbitration or class-action waiver
Popularity1.1k stars, 2.7k PyPI/wk555 stars, 13k PyPI/wk

Verdicts

Decider

An Apache-2.0 decision model family with a dated changelog, passing CI, 21 package releases since 22 September 2026 and model cards that list measured regressions. One person maintains it, the local server has no authentication option, states over 32,768 tokens are cut without an error, and no security policy is published.

GLiClass

An Apache-2.0 classifier that scores a whole label set in one encoder pass on the owner's hardware, with single-label, multi-label, hierarchical and few-shot modes. It returns label scores with no calibration claim, the bundled server has no authentication, and the last three test runs on the main branch, on 24 September 2026, failed.

Before you call either

Decider

  1. Pin weights by Hub tag (v10, v2) when results must repeat. The main branch of each model repository changes with new versions
  2. Read x_p_max for the top probability. Since 1.3.0 confidence on choice and score answers follows TypeSafe's rescaled definition, not the top probability
  3. Keep the server on 127.0.0.1 or put an authenticating proxy in front. It has no key option
  4. Count state tokens before sending. Over 32,768 the state is cut silently, and /decide caps context at 1,536 tokens
  5. Split multi-step arithmetic or multi-hop judgements into several questions, and don't write long rules into a question. The README says both fail

GLiClass

  1. Pass --host 127.0.0.1 to python -m gliclass.serve, or put the port behind your own gateway. The server checks no credential
  2. On a machine without a GPU add --device cpu --dtype float32 --num-gpus-per-replica 0. The default configuration expects CUDA
  3. Send one text a request to POST /gliclass. An array in texts is cut to its first item without an error
  4. Set multi_label to false for one label from a set. The default scores each label independently, so scores do not sum to 1
  5. Keep text plus labels under the pipeline's 1,024-token max_length, or use ZeroShotClassificationWithChunkingPipeline. Longer input is truncated silently

Questions

Which is better for AI agents, Decider or GLiClass?

Decider scores 69.5 (B) on agent readiness against GLiClass's 49.9 (D), and leads in 4 of 7 scored categories. GLiClass leads on transparency & trust.

Do Decider and GLiClass need an API key?

Neither needs a key.

Can an agent call Decider and GLiClass without installing anything?

No hosted endpoint is listed for Decider. No hosted endpoint is listed for GLiClass.

Are Decider and GLiClass open source?

Yes. Decider is open source (Apache-2.0 (code and weights)). GLiClass is open source (Apache-2.0 (library and the model weights we checked)).

Other comparisons with Decider or GLiClass

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.