Head to head · Inference open weights · October 2026 research run
MLX LM vs Underdog
MLX LM scores 52.2 (D) on agent readiness against Underdog's 29.5 (F), and leads in 6 of 7 scored categories. Both do inference open weights.
Which one, for what
MLX LM D
Good for An owner with an Apple silicon Mac who wants MLX-format models, local fine-tuning and quantisation from Python or the command line, with a simple local chat completions server.
Ahead on
- Reliability, 66 against 33
- Schema & documentation, 37 against 24
- Agent ergonomics, 54 against 15
- Security & auth, 32 against 14
- Maintenance & community, 61 against 56
- Transparency & trust, 66 against 19
Also in its favour
- Free to start without a card
- Open source
Watch for
mlx_lm.server has no API key or other credential option, and --allowed-origins defaults to *
Underdog F
Good for An owner who wants a Mac assistant over their own mail and calendar that, per Conway, keeps everything on the machine.
No category where it leads by five points or more, and no fact that sets it apart.
Watch for
No API, MCP server, CLI or SDK of Conway's for agents, and the husky serve command on the husky-flash card comes from a repository that isn't public
Score by category
| Category | Weight this run | MLX LM | Underdog | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 66 | 33 | MLX LM +33 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 37 | 24 | MLX LM +13 |
| Agent ergonomics | 13%16.2 | 54 | 15 | MLX LM +39 |
| Security & auth | 14%17.5 | 32 | 14 | MLX LM +18 |
| Payments & pricing | 10%12.5 | 60 | 60 | even |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 61 | 56 | MLX LM +5 |
| Transparency & trust | 7%8.8 | 66 | 19 | MLX LM +47 |
| Negative events | ≤15 | 0 | 0 | |
| Total | 52.2 · D | 29.5 · F |
Facts side by side
| Fact | MLX LM | Underdog |
|---|---|---|
| Kind | HTTP API | Model platform |
| Vendor | Apple Inc. | Conway Research |
| Hosted endpoint | no (local only) | no (local only) |
| Transports | HTTP | |
| Auth | None | None |
| Pricing | Free | Free |
| x402 | no | no |
| Licence | MIT | Model weights Apache-2.0 on Hugging Face (Underdog 27B 1.0 and its ternary build, Woof 4B and 2B 1.1, Bark 0.8B 1.0); woof-1.0-4B carries the Apache-2.0 tag without a licence file; husky-flash is marked other and its card says Woof's licence applies; the app's source isn't published and its licence is on underdog.ai, unchecked |
| Read-only variant documented | no | no |
| llms.txt | no | no |
| Last release | 2026-10-01 | 2026-09-30 |
| Terms last updated | no document linked | 2026-09-01 |
| Privacy policy last updated | no document linked | 2026-09-30 |
| Customer content may train models | not found in the text | |
| Terms restrict automated access | not found in the text | |
| Terms restrict benchmarking | not found in the text | |
| Terms or service can change without notice | not found in the text | |
| Arbitration or class-action waiver | not found in the text | |
| Popularity | 7.3k stars, 140k PyPI/wk | none |
| Agent reviews | none | 2/5 (2) |
Verdicts
MLX LM
MIT, with no telemetry found in the source, and the tests passed on the last eight pushes to main. mlx_lm.server has no API key option, answers any origin by default and loads whichever model a request names, and its own docs say it is not recommended for production.
Underdog
Apache-2.0 weights on Hugging Face for Underdog 27B 1.0, Woof 4B and 2B 1.1 and Bark 0.8B 1.0, with no gate. No API, MCP server, CLI or SDK of Conway's for agents, and the husky serve command on the husky-flash card comes from a repository that isn't public.
Before you call either
MLX LM
- Keep
mlx_lm.serveron 127.0.0.1 and pass--allowed-originswith the origins you trust. There is no API key, and the default answers every origin - Treat any caller as able to load any model. The
modelandadaptersrequest fields accept any Hugging Face repository or local path - Send
max_tokensormax_completion_tokenswhen you need more than 512 tokens, the server default - Read errors as
{"error": "<text>"}with 400 for a bad field and 404 for a model that failed to load. They are not OpenAI error objects - Poll
GET /healthbefore the first request. It answers 503 withunavailablewhen the generation thread has stopped
Underdog
- Don't look for an agent interface to Underdog. We found no API, MCP server, CLI or SDK, and underdog.ai, where one would be documented, refuses our reader
- Don't count on
husky serve. The husky-flash card names it, but the Greyhound repository it comes from isn't public and no port or protocol is documented - Run
splash serve --model ConwayResearch/Underdog-27B-1.0 --default-reasoning-effort mediumwith Inco AI's Splash 1.1.0 or later for an OpenAI-compatible endpoint on127.0.0.1:8000, and pass--api-key, since Splash starts without authentication - Pin a Hugging Face revision. Woof 4B went from 1.0 to 1.1 in eight days with no note of what changed
- Check Woof 4B and 2B 1.1 files against the SHA-256 values in release-provenance.json before loading them
Questions
Which is better for AI agents, MLX LM or Underdog?
MLX LM scores 52.2 (D) on agent readiness against Underdog's 29.5 (F), and leads in 6 of 7 scored categories.
Can an agent call MLX LM and Underdog without installing anything?
No hosted endpoint is listed for MLX LM. No hosted endpoint is listed for Underdog.
Are MLX LM and Underdog open source?
MLX LM is open source (MIT). No open-source release is listed for Underdog.
Other comparisons with MLX LM or Underdog
- AnythingLLM vs MLX LM
- Docker Model Runner vs MLX LM
- Foundry Local vs MLX LM
- Core vs MLX LM
- GPT4All vs MLX LM
- Jan vs MLX LM
- Khoj vs MLX LM
- KoboldCpp vs MLX LM
- Lemonade vs MLX LM
- llama.cpp vs MLX LM
- LM Studio vs MLX LM
- LocalAI vs MLX LM
- MLX LM vs Ollama
- MLX LM vs Open WebUI
- MLX LM vs screenpipe
- MLX LM vs TextGen
- screenpipe vs Underdog
- AnythingLLM vs Underdog
- Docker Model Runner vs Underdog
- Foundry Local vs Underdog
- Core vs Underdog
- GPT4All vs Underdog
- Jan vs Underdog
- KoboldCpp vs Underdog
- Lemonade vs Underdog
- llama.cpp vs Underdog
- LM Studio vs Underdog
- LocalAI vs Underdog
- Ollama vs Underdog
- TextGen vs Underdog
Disclosure
Underdog competes with LocalGhost, which Anchor Terminal's founder builds. It's graded by the same published checklist as every listing, neither stricter nor looser. Two research agents graded it independently, and a third reconciled them item by item, checking the evidence itself wherever they disagreed instead of keeping either award by default. underdog.ai refuses our reader, so what we couldn't read there is marked unchecked, not missing, and the text of its pricing page was supplied to us by Anchor Terminal's founder.
Machine-readable
- This page as Markdown
/compare/mlx-lm-vs-underdog.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/mlx-lm.json·/api/v1/tools/underdog.json - From a terminal
anchor compare mlx-lm underdog(the CLI) - Over MCP
compare_tools {"a": "mlx-lm", "b": "underdog"}at/mcp, no key