Head to head · LLM inference · October 2026 research run

Cloudflare Workers AI vs SiliconFlow

Cloudflare Workers AI scores 68.3 (B) on agent readiness against SiliconFlow's 46.7 (D), and leads in every scored category. Both do llm inference.

Best model APIs and inference for AI agents · All 136 models comparisons

Which one, for what

Cloudflare Workers AI B

Good for Agents that already run on Cloudflare Workers, or that want open-weight chat, embedding, reranking, speech and image models behind one token with a free daily allowance.

Ahead on

  • Reliability, 66 against 33
  • Schema & documentation, 80 against 65
  • Agent ergonomics, 75 against 60
  • Security & auth, 77 against 39
  • Payments & pricing, 52 against 35
  • Maintenance & community, 72 against 47
  • Transparency & trust, 76 against 51

Watch for

No SLA for Workers AI was found in the docs or the agreements read

SiliconFlow D

Good for Agents that want recent open-weight chat models, plus image, video and speech, behind one OpenAI-style key at low per-token prices, and can live without a status page or SLA.

Also in its favour

  • No incidents deducted, where Cloudflare Workers AI loses 3 points for them

Watch for

No status page, incident history or SLA was found, and the terms disclaim any uptime or availability commitment

Score by category

CategoryWeight this runCloudflare Workers AISiliconFlowEdge
Reliability16%206633Cloudflare Workers AI +33
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.28065Cloudflare Workers AI +15
Agent ergonomics13%16.27560Cloudflare Workers AI +15
Security & auth14%17.57739Cloudflare Workers AI +38
Payments & pricing10%12.55235Cloudflare Workers AI +17
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.87247Cloudflare Workers AI +25
Transparency & trust7%8.87651Cloudflare Workers AI +25
Negative events≤15-30
Total68.3 · B46.7 · D

Facts side by side

FactCloudflare Workers AISiliconFlow
KindModel APIModel API
VendorCloudflare, Inc.SiliconFlow Labs Pte. Ltd.
Hosted endpointhttps://api.cloudflare.com/client/v4/accounts/{account_id}/aihttps://api.siliconflow.com/v1
TransportsHTTPHTTP
AuthAPI keyAPI key
PricingFreemiumPay per use
x402payer tooling onlyno
LicenceProprietary service under Cloudflare's Self-Serve Subscription Agreement. Each hosted model carries its own open-weight licence, linked from its model page. The cloudflare SDKs are Apache-2.0, workers-ai-provider is MIT and the OpenAPI repository is BSD-3-ClauseProprietary service under the SiliconFlow Terms of Use
Read-only variant documentednono
llms.txtyesyes
Last release2026-10-012026-09-14
Terms last updated2025-09-12no date given
Privacy policy last updatedno date givenno date given
Customer content may train modelsnot found in the textnot found in the text
Terms restrict automated accessyesyes
Terms restrict benchmarkingnot found in the textyes
Terms or service can change without noticeyesyes
Arbitration or class-action waiveryesyes
Popularity380k npm/wknone

Verdicts

Cloudflare Workers AI

Per-model prices, rate limits and JSON Schemas are public, 10,000 neurons a day are free, and x402 payment is in beta on /ai/run for four models. No SLA was found, three models moved to paid-only access on 28 July 2026 with no notice, and several docs pages still use a model retired in May.

SiliconFlow

Per-token prices for every listed model are public, and one key reaches chat, embeddings, reranking, image, video and speech through OpenAI-style and Anthropic-style routes. No status page, SLA, security page or official SDK was found, release notes stop at 11 June 2026, and several model removals are dated the same day as their notice.

Before you call either

Cloudflare Workers AI

  1. Check the model page before calling. Seven models need Workers Paid or AI Gateway credits and return 403 with code 5035 on the Free plan
  2. Read the internal code on a 429. 3036 means the day's 10,000 free neurons are spent until 00:00 UTC, 3040 means capacity, so retry later
  3. Set options.rejectIfBusy to fail fast, or add ?queueRequest=true for the batch route when the answer can wait
  4. Send the same x-session-affinity value on every turn of a session to reach the prefix cache and the cached-input price
  5. Keep paid frontier models under 20 requests a minute per model, or 50 with prepaid AI Gateway credits

SiliconFlow

  1. Set the OpenAI client's base URL to https://api.siliconflow.com/v1, or the Anthropic client's to https://api.siliconflow.com/, and send the key as a Bearer token
  2. Take model ids from the model pages or GET /v1/models, not from the OpenAPI enum or the function calling guide, which list removed models
  3. Check the model field of each response. On 11 June 2026 traffic for GLM-5 and Kimi-K2.5 was routed to successor models
  4. Use response_format of json_object only where the model page says JSON Mode is supported, and keep max_tokens about 10,000 below the context length
  5. Set the Claude Code environment variables by hand. The automated route pipes a script from an Amazon S3 bucket into bash

Questions

Which is better for AI agents, Cloudflare Workers AI or SiliconFlow?

Cloudflare Workers AI scores 68.3 (B) on agent readiness against SiliconFlow's 46.7 (D), and leads in every scored category.

Do Cloudflare Workers AI and SiliconFlow need an API key?

Both need an API key.

Can an agent call Cloudflare Workers AI and SiliconFlow without installing anything?

Yes. Cloudflare Workers AI has a hosted endpoint at https://api.cloudflare.com/client/v4/accounts/{account_id}/ai and SiliconFlow at https://api.siliconflow.com/v1.

Other comparisons with Cloudflare Workers AI or SiliconFlow

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.