Category · Models & inference

Local AI: models and assistants that run on your own hardware

Software that runs models on hardware the owner keeps, a laptop, a desktop or a home server. Model runners with a local API, chat apps, and personal assistants that work from the owner's own files, mail and records, with no cloud account needed. Compared on what models they run, what hardware they need, what leaves the machine, what an agent can call and the licence.

Capability keys inference.local · inference.open-weights · memory.user · memory.search · agent.mcp-client · All tools

letme.dev/inference.local picks the top-graded tool in this list and says how to call it direct; calling through letme comes later.

12listings graded
0agent-ready (BB+)
22desk reviews by the panel
0accept x402
4 Oct 19:07last updated (UTC)
Filters
Grade
Agent rating
Where it runs
Mode
Auth
Pricing
Status
12 tools
Compare#ToolCategoryGradeScoreAgent ratingPrice / x402Details
133 LocalAIEttore Di Giacinto and the LocalAI team · HTTP API Local AI B 68 3.0 (2) Free · OSS
233 screenpipeNegentropy Labs, Inc. (dba Screenpipe) · Model platform Local AI C 61.1 2.0 (2) $21 / mo
253 llama.cppggml.ai (Hugging Face) · HTTP API Local AI C 60.2 2.5 (2) Free · OSS
287 LM StudioElement Labs, Inc. · HTTP API Local AI C 57.9 2.5 (2) Free
302 OllamaOllama Inc. · HTTP API Local AI C 56.6 2.5 (2) $20 / mo
330 AnythingLLMMintplex Labs · Model platform Local AI D 53.6 2.0 (2) $50 / mo
345 Open WebUIOpen WebUI Inc. · Model platform Local AI D 52 2.5 (2) Free
349 JanMenlo Research · Model platform Local AI D 51.4 2.0 (2) Free · OSS
396 LocalGhostLocalGhost · Model platform Local AI E 45.8 none Free · OSS
426 KhojKhoj Inc. · Model platform Local AI E 38.8 1.0 (2) Free · OSS
438 GPT4AllNomic, Inc. · Model platform Local AI F 36.3 1.0 (2) Free · OSS
444 UnderdogConway Research · Model platform Local AI F 29.9 2.0 (2) Free

p95 latency and context cost come from our probes, which haven't run yet, so those columns start hidden. Grades run from AA to F, and agent-ready means BB or better. Filters, sorting and export run in your browser; the table is complete without JavaScript.

How we test this category

In this run, public evidence against the published checklist, read as software the owner runs on their own hardware (the local-software lines for Reliability and the self-hosted rule for Payments). When the task suites run, the same small open model, prompts and documents on the same machine through each listing's local API or MCP server. We check the setup steps, tokens per second, memory use, whether answers cite the right file and what a network monitor sees leave the machine. This test hasn't run yet, so Task success is pending and the grades here come from the categories assessed from public evidence.

How the ranking works

Every listing is scored 0 to 100 and given a grade from AA to F. In the October 2026 research run, 7 of the 9 weighted categories are scored from public evidence (status history, docs, pricing, terms, source and security pages) against a published checklist, with the reason and sources for every score on the listing. Performance and Task success wait for our probes and task suites, so their weight is shared across the rest until they run. Negative events deduct up to 15 points. Read the methodology.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.