Head to head · Local inference · October 2026 research run
Foundry Local vs MLX LM
Foundry Local scores 60.5 (C) on agent readiness against MLX LM's 52.2 (D), and leads in 6 of 7 scored categories. Both do local inference.
Which one, for what
Good for An application that ships a model to end users' Windows, macOS or Linux devices and wants NPU and GPU variants chosen automatically, especially on Windows.
Ahead on
- Schema & documentation, 53 against 37
- Agent ergonomics, 61 against 54
- Security & auth, 39 against 32
- Maintenance & community, 88 against 61
- Transparency & trust, 72 against 66
Watch for
The local server takes no credential, and its routes include model load and unload and POST /shutdown
MLX LM D
Good for An owner with an Apple silicon Mac who wants MLX-format models, local fine-tuning and quantisation from Python or the command line, with a simple local chat completions server.
No category where it leads by five points or more, and no fact that sets it apart.
Watch for
mlx_lm.server has no API key or other credential option, and --allowed-origins defaults to *
Score by category
| Category | Weight this run | Foundry Local | MLX LM | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 68 | 66 | Foundry Local +2 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 53 | 37 | Foundry Local +16 |
| Agent ergonomics | 13%16.2 | 61 | 54 | Foundry Local +7 |
| Security & auth | 14%17.5 | 39 | 32 | Foundry Local +7 |
| Payments & pricing | 10%12.5 | 60 | 60 | even |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 88 | 61 | Foundry Local +27 |
| Transparency & trust | 7%8.8 | 72 | 66 | Foundry Local +6 |
| Negative events | ≤15 | 0 | 0 | |
| Total | 60.5 · C | 52.2 · D |
Facts side by side
| Fact | Foundry Local | MLX LM |
|---|---|---|
| Kind | SDK + MCP | HTTP API |
| Vendor | Microsoft | Apple Inc. |
| Hosted endpoint | no (local only) | no (local only) |
| Transports | HTTP | HTTP |
| Auth | None | None |
| Pricing | Free | Free |
| x402 | no | no |
| Licence | MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own | MIT |
| Read-only variant documented | no | no |
| llms.txt | no | no |
| Last release | 2026-09-29 | 2026-10-01 |
| Terms last updated | no document linked | no document linked |
| Privacy policy last updated | no document linked | no document linked |
| Customer content may train models | ||
| Terms restrict automated access | ||
| Terms restrict benchmarking | ||
| Terms or service can change without notice | ||
| Arbitration or class-action waiver | ||
| Popularity | 2.6k stars, 105k npm/wk, 47k PyPI/wk | 7.3k stars, 140k PyPI/wk |
Verdicts
Foundry Local
The SDK and its native runtime are MIT, at 2.1.0 on four registries, and pick a CPU, GPU or NPU model variant automatically. The optional local server has no credential, its current routes aren't in the published REST reference, and telemetry is on by default with an opt-out.
MLX LM
MIT, with no telemetry found in the source, and the tests passed on the last eight pushes to main. mlx_lm.server has no API key option, answers any origin by default and loads whichever model a request names, and its own docs say it is not recommended for production.
Before you call either
Foundry Local
- Read the server URL from
manager.urls[0]orfoundry server status. The port is dynamic unless the owner setsweb.urlsorfoundry server start --port - Send the model ID that
GET /v1/modelsreturns, not the alias. The alias resolves to a hardware-specific variant - Check
supportsToolCallingbefore sending tools. Support differs by variant, and issue #1183 reports Qwen tool calling failing on QNN - Set your own request timeout. Inference has no built-in one, and cancellation takes effect only after the current generation step
- Ask the owner to set
ORT_TELEMETRY_DISABLED=1ordisableNonessentialTelemetrybefore the manager is created if telemetry must be off
MLX LM
- Keep
mlx_lm.serveron 127.0.0.1 and pass--allowed-originswith the origins you trust. There is no API key, and the default answers every origin - Treat any caller as able to load any model. The
modelandadaptersrequest fields accept any Hugging Face repository or local path - Send
max_tokensormax_completion_tokenswhen you need more than 512 tokens, the server default - Read errors as
{"error": "<text>"}with 400 for a bad field and 404 for a model that failed to load. They are not OpenAI error objects - Poll
GET /healthbefore the first request. It answers 503 withunavailablewhen the generation thread has stopped
Questions
Which is better for AI agents, Foundry Local or MLX LM?
Foundry Local scores 60.5 (C) on agent readiness against MLX LM's 52.2 (D), and leads in 6 of 7 scored categories.
Can an agent call Foundry Local and MLX LM without installing anything?
No hosted endpoint is listed for Foundry Local. No hosted endpoint is listed for MLX LM.
Are Foundry Local and MLX LM open source?
Yes. Foundry Local is open source (MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own). MLX LM is open source (MIT).
Other comparisons with Foundry Local or MLX LM
- AnythingLLM vs Foundry Local
- AnythingLLM vs MLX LM
- Docker Model Runner vs Foundry Local
- Docker Model Runner vs MLX LM
- Foundry Local vs Core
- Foundry Local vs GPT4All
- Foundry Local vs Jan
- Foundry Local vs Khoj
- Foundry Local vs KoboldCpp
- Foundry Local vs Lemonade
- Foundry Local vs llama.cpp
- Foundry Local vs LM Studio
- Foundry Local vs LocalAI
- Foundry Local vs Ollama
- Foundry Local vs Open WebUI
- Foundry Local vs screenpipe
- Foundry Local vs TextGen
- Core vs MLX LM
- GPT4All vs MLX LM
- Jan vs MLX LM
- Khoj vs MLX LM
- KoboldCpp vs MLX LM
- Lemonade vs MLX LM
- llama.cpp vs MLX LM
- LM Studio vs MLX LM
- LocalAI vs MLX LM
- MLX LM vs Ollama
- MLX LM vs Open WebUI
- MLX LM vs screenpipe
- MLX LM vs TextGen
- Foundry Local vs Underdog
- MLX LM vs Underdog
Machine-readable
- This page as Markdown
/compare/foundry-local-vs-mlx-lm.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/foundry-local.json·/api/v1/tools/mlx-lm.json - From a terminal
anchor compare foundry-local mlx-lm(the CLI) - Over MCP
compare_tools {"a": "foundry-local", "b": "mlx-lm"}at/mcp, no key