Head to head · Local inference · October 2026 research run

Foundry Local vs TextGen

Foundry Local scores 60.5 (C) on agent readiness against TextGen's 45.1 (E), and leads in 4 of 7 scored categories. TextGen leads on schema & documentation. Both do local inference.

Which one, for what

Foundry Local C

Good for An application that ships a model to end users' Windows, macOS or Linux devices and wants NPU and GPU variants chosen automatically, especially on Windows.

Ahead on

  • Reliability, 68 against 51
  • Agent ergonomics, 61 against 50
  • Maintenance & community, 88 against 24
  • Transparency & trust, 72 against 49

Also in its favour

  • No key needed to call it
  • No incidents deducted, where TextGen loses 4 points for them

Watch for

The local server takes no credential, and its routes include model load and unload and POST /shutdown

TextGen E

Good for A person who wants one app for several backends (llama.cpp, ExLlamaV3, Transformers) with an OpenAI and Anthropic-compatible endpoint, LoRA training and image generation.

Ahead on

  • Schema & documentation, 61 against 53

Watch for

No release since v4.9 on 20 May 2026 and no code commit on main or dev since 31 May 2026

Score by category

CategoryWeight this runFoundry LocalTextGenEdge
Reliability16%206851Foundry Local +17
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.25361TextGen +8
Agent ergonomics13%16.26150Foundry Local +11
Security & auth14%17.53940TextGen +1
Payments & pricing10%12.56060even
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88824Foundry Local +64
Transparency & trust7%8.87249Foundry Local +23
Negative events≤150-4
Total60.5 · C45.1 · E

Facts side by side

FactFoundry LocalTextGen
KindSDK + MCPModel platform
VendorMicrosoftoobabooga
Hosted endpointno (local only)no (local only)
TransportsHTTPHTTP
AuthNoneAPI key
PricingFreeFree
x402nono
LicenceMIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its ownAGPL-3.0
Read-only variant documentednono
llms.txtnono
Last release2026-09-292026-05-20
Terms last updatedno document linkedno document linked
Privacy policy last updatedno document linkedno document linked
Customer content may train models
Terms restrict automated access
Terms restrict benchmarking
Terms or service can change without notice
Arbitration or class-action waiver
Popularity2.6k stars, 105k npm/wk, 47k PyPI/wk48k stars

Verdicts

Foundry Local

The SDK and its native runtime are MIT, at 2.1.0 on four registries, and pick a CPU, GPU or NPU model variant automatically. The optional local server has no credential, its current routes aren't in the published REST reference, and telemetry is on by default with an opt-out.

TextGen

AGPL-3.0 with no telemetry, and a local API on 127.0.0.1:5000 that checks the Host header, limits CORS to localhost and separates an admin key from the caller's key. No release since v4.9 on 20 May 2026, no code commits since 31 May, no test suite, and ten security advisories in the year, all fixed.

Before you call either

Foundry Local

  1. Read the server URL from manager.urls[0] or foundry server status. The port is dynamic unless the owner sets web.urls or foundry server start --port
  2. Send the model ID that GET /v1/models returns, not the alias. The alias resolves to a hardware-specific variant
  3. Check supportsToolCalling before sending tools. Support differs by variant, and issue #1183 reports Qwen tool calling failing on QNN
  4. Set your own request timeout. Inference has no built-in one, and cancellation takes effect only after the current generation step
  5. Ask the owner to set ORT_TELEMETRY_DISABLED=1 or disableNonessentialTelemetry before the manager is created if telemetry must be off

TextGen

  1. Ask the owner to launch with --api. Nothing listens on port 5000 without it, and a model must be loaded first
  2. Call http://127.0.0.1:5000/v1. Any Host header other than localhost or 127.0.0.1 gets 400 Invalid host header unless --listen is set
  3. Send the key as Authorization: Bearer on OpenAI routes and as x-api-key on /v1/messages. Model loading needs the admin key
  4. Read http://127.0.0.1:5000/docs or modules/api/typing.py for parameters. max_tokens defaults to 512 on chat completions
  5. Run tool calls yourself. The API returns finish_reason: "tool_calls" and executes nothing on the server
  6. Use the repository name oobabooga/textgen. The old text-generation-webui URL redirects

Questions

Which is better for AI agents, Foundry Local or TextGen?

Foundry Local scores 60.5 (C) on agent readiness against TextGen's 45.1 (E), and leads in 4 of 7 scored categories. TextGen leads on schema & documentation.

Are Foundry Local and TextGen open source?

Yes. Foundry Local is open source (MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own). TextGen is open source (AGPL-3.0).

Other comparisons with Foundry Local or TextGen

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.