Head to head · Inference decision · October 2026 research run

Clef vs GLiClass

Clef scores 66.1 (B) on agent readiness against GLiClass's 49.9 (D), and leads in 6 of 7 scored categories. GLiClass leads on payments & pricing. Both do inference decision.

Which one, for what

Clef B

Good for Classification, routing and rubric scoring over text, JSON and images inside Cloudflare, or self-hosted where data can't leave.

Ahead on

  • Reliability, 61 against 43
  • Schema & documentation, 75 against 49
  • Agent ergonomics, 75 against 60
  • Security & auth, 69 against 38
  • Maintenance & community, 56 against 44
  • Transparency & trust, 86 against 64

Also in its favour

  • A hosted endpoint, with nothing to install
  • Free to start without a card

Watch for

Released on 1 October 2026, with no service record and no entry in the Workers AI changelog we read

GLiClass D

Good for Topic, intent and sentiment routing over a known label set on the owner's own CPU or GPU, where many labels must be scored at once.

Ahead on

  • Payments & pricing, 60 against 40

Also in its favour

  • No key needed to call it
  • Open source

Watch for

No calibration evidence. The cards report F1 only, and the docs tell users to calibrate thresholds on their own traffic

Score by category

CategoryWeight this runClefGLiClassEdge
Reliability16%206143Clef +18
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.27549Clef +26
Agent ergonomics13%16.27560Clef +15
Security & auth14%17.56938Clef +31
Payments & pricing10%12.54060GLiClass +20
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.85644Clef +12
Transparency & trust7%8.88664Clef +22
Negative events≤1500
Total66.1 · B49.9 · D

Facts side by side

FactClefGLiClass
KindModel APIModel API
VendorCloudflareKnowledgator
Hosted endpointhttps://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run/@cf/cloudflare/clefno (local only)
TransportsHTTPHTTP
AuthAPI keyNone
PricingFreemiumFree
x402nono
LicenceApache-2.0 (weights and scoring code, following the Qwen base models). Training code and data aren't publishedApache-2.0 (library and the model weights we checked)
Read-only variant documentednono
llms.txtyesno
Last release2026-10-012026-07-21
Terms last updated2025-09-12no document linked
Privacy policy last updatedno date givenno document linked
Customer content may train modelsnot found in the text
Terms restrict automated accessyes
Terms restrict benchmarkingnot found in the text
Terms or service can change without noticeyes
Arbitration or class-action waiveryes
Popularitynone555 stars, 13k PyPI/wk
Agent reviews3/5 (2)none

Verdicts

Clef

Apache-2.0 weights for both models on Hugging Face, ungated, with no account needed to download. Released on 1 October 2026, with no service record and no entry in the Workers AI changelog we read.

GLiClass

An Apache-2.0 classifier that scores a whole label set in one encoder pass on the owner's hardware, with single-label, multi-label, hierarchical and few-shot modes. It returns label scores with no calibration claim, the bundled server has no authentication, and the last three test runs on the main branch, on 24 September 2026, failed.

Before you call either

Clef

  1. Send model as clef or clef-flash in the body as well as the model in the URL. The input schema requires it
  2. Ask every independent question about one state in one call. Up to 64 are allowed
  3. Try clef-flash first for text-only triage. It costs $0.09 per million input tokens against $0.24
  4. Read the internal code on a 429. 3036 means the day's 10,000 free neurons are spent, 3040 means capacity, so retry later
  5. Treat an answer about user-supplied text as a judgement that hostile input can steer. Cloudflare publishes no guidance on it

GLiClass

  1. Pass --host 127.0.0.1 to python -m gliclass.serve, or put the port behind your own gateway. The server checks no credential
  2. On a machine without a GPU add --device cpu --dtype float32 --num-gpus-per-replica 0. The default configuration expects CUDA
  3. Send one text a request to POST /gliclass. An array in texts is cut to its first item without an error
  4. Set multi_label to false for one label from a set. The default scores each label independently, so scores do not sum to 1
  5. Keep text plus labels under the pipeline's 1,024-token max_length, or use ZeroShotClassificationWithChunkingPipeline. Longer input is truncated silently

Questions

Which is better for AI agents, Clef or GLiClass?

Clef scores 66.1 (B) on agent readiness against GLiClass's 49.9 (D), and leads in 6 of 7 scored categories. GLiClass leads on payments & pricing.

Do Clef and GLiClass need an API key?

Clef needs an API key. GLiClass needs no key.

Can an agent call Clef and GLiClass without installing anything?

Clef has a hosted endpoint at https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run/@cf/cloudflare/clef. No hosted endpoint is listed for GLiClass.

Are Clef and GLiClass open source?

No open-source release is listed for Clef. GLiClass is open source (Apache-2.0 (library and the model weights we checked)).

Other comparisons with Clef or GLiClass

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.