# Foundry Local > Microsoft's on-device model runtime, built on ONNX Runtime. Applications embed it through SDKs for C#, JavaScript, Python and Rust, and it can start an optional OpenAI-compatible server on localhost. A preview CLI is also available. - Canonical: https://www.anchorterminal.com/tools/foundry-local - Markdown: https://www.anchorterminal.com/tools/foundry-local.md (~8,200 tokens) - Slim: https://www.anchorterminal.com/tools/foundry-local.min.md (~1,780 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/tools/foundry-local.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-09 ## Overview **Grade C · 60.5/100 · rank #461 of 842 · #4 in Local AI · not agent-ready · confidence medium** More from Microsoft, listed separately because each is its own product: [Microsoft Foundry fine-tuning (Azure OpenAI)](https://www.anchorterminal.com/tools/azure-foundry-fine-tuning.md) (Fine-tuning), [Azure AI Content Safety (Prompt Shields)](https://www.anchorterminal.com/tools/azure-ai-content-safety.md) (Guardrails & safety filters), [Azure AI Speech speech-to-text](https://www.anchorterminal.com/tools/azure-speech-to-text.md) (Speech-to-text), [Azure AI Speech text-to-speech](https://www.anchorterminal.com/tools/azure-text-to-speech.md) (Text-to-speech), [Microsoft Agent Framework](https://www.anchorterminal.com/tools/microsoft-agent-framework.md) (Agent frameworks & SDKs), [Microsoft Execution Containers](https://www.anchorterminal.com/tools/microsoft-execution-containers.md) (Code execution sandboxes), [Microsoft Entra Agent ID](https://www.anchorterminal.com/tools/microsoft-entra-agent-id.md) (Agent auth & delegated access), [Azure Key Vault](https://www.anchorterminal.com/tools/azure-key-vault.md) (Secrets & credential vaults), [Azure Document Intelligence](https://www.anchorterminal.com/tools/azure-document-intelligence.md) (Document parsing & extraction), [Azure DevOps MCP Server](https://www.anchorterminal.com/tools/azure-devops-mcp.md) (Code & developer platforms), [Microsoft Learn MCP Server](https://www.anchorterminal.com/tools/microsoft-learn-mcp.md) (Code & developer platforms), [Playwright MCP](https://www.anchorterminal.com/tools/playwright-mcp.md) (Browser automation), [Azure MCP Server](https://www.anchorterminal.com/tools/azure-mcp.md) (Cloud & infrastructure), [Azure Maps](https://www.anchorterminal.com/tools/azure-maps.md) (Maps, geocoding & places), [Azure Translator](https://www.anchorterminal.com/tools/azure-translator.md) (Translation), [Microsoft Graph Calendar API](https://www.anchorterminal.com/tools/microsoft-graph-calendar.md) (Calendars & scheduling), [Azure Blob Storage](https://www.anchorterminal.com/tools/azure-blob-storage.md) (File storage & sharing), [OneDrive and SharePoint files (Microsoft Graph)](https://www.anchorterminal.com/tools/onedrive-sharepoint.md) (File storage & sharing), [Microsoft Teams (Microsoft Graph)](https://www.anchorterminal.com/tools/microsoft-teams.md) (Work & productivity), [Microsoft Dynamics 365 Sales](https://www.anchorterminal.com/tools/dynamics-365-sales.md) (CRM & customer platforms), [Microsoft Power Automate](https://www.anchorterminal.com/tools/power-automate.md) (Workflow automation), [Microsoft Advertising API](https://www.anchorterminal.com/tools/microsoft-advertising-api.md) (Advertising & campaign operations), [Microsoft Excel (Microsoft Graph workbook API)](https://www.anchorterminal.com/tools/microsoft-excel-graph.md) (Spreadsheets & operational tables), [Outlook Mail (Microsoft Graph)](https://www.anchorterminal.com/tools/outlook-mail-graph.md) (Mailbox access). ## Assessment The SDK and its native runtime are MIT, at 2.1.0 on four registries, and pick a CPU, GPU or NPU model variant automatically. The optional local server has no credential, its current routes aren't in the published REST reference, and telemetry is on by default with an opt-out. ## Facts | Field | Value | | --- | --- | | Vendor | Microsoft (https://www.foundrylocal.ai) | | Kind | SDK + MCP | | Category | Local AI (https://www.anchorterminal.com/categories/local-ai) | | Transport | HTTP | | Auth | None · No authentication. The SDK runs in the application's own process, and the optional local server takes no key or token. It binds to 127.0.0.1 on a dynamic port unless the owner configures `web.urls` or starts the CLI daemon with `--port`. No account or Azure subscription is needed. | | Pricing | Free (Free · OSS) · Free, with no account, card or Azure subscription. The SDK is MIT and the CLI is a free download under Microsoft's licence terms. Microsoft states there are no per-token costs. Foundry Local on Azure Local is a separate product for servers and is not covered here (checked 2026-10-08). | | x402 | No · No x402, MPP or L402 in the docs, the README or the SDK source (checked 2026-10-08). | | Licence | MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own | | Packages | npm: `foundry-local-sdk`; pypi: `foundry-local-sdk`; nuget: `Microsoft.AI.Foundry.Local`; cargo: `foundry-local-sdk` | | Source | https://github.com/microsoft/Foundry-Local | | Docs | https://learn.microsoft.com/en-us/azure/foundry-local/ | | llms.txt | not found | | Last release | 2026-09-29 | | GitHub stars | 2,600 (as of 2026-10-08) | | npm downloads / week | 105,372 | | PyPI downloads / week | 47,344 | | Interfaces | SDKs for C#, JavaScript, Python and Rust, with C++, a C ABI and Java in the v2 source. Optional OpenAI-compatible server started by `start_web_service()` or `foundry server start`. `foundry` CLI in public preview | | Local server | http://127.0.0.1 on a dynamic port by default. `/v1/chat/completions`, `/v1/responses` (stored, with `previous_response_id`), `/v1/embeddings`, `/v1/audio/transcriptions`, `/v1/models`, `/models/load/{name}`, `/models/unload/{name}`, `/models/loaded`, `/status`, `POST /shutdown`, per the v2 source | | Credentials | None. The docs tell OpenAI clients to send any placeholder key and set Open WebUI's auth to None | | Engine | ONNX Runtime 1.30.0 and ONNX Runtime GenAI 0.17.1 in v2. Models are ONNX, from the Foundry catalogue or registered by the owner (bring your own model, 2.1.0) | | Hardware | CPU everywhere. WebGPU through Dawn on Windows, Linux and macOS (Metal on Apple silicon). NVIDIA CUDA on Windows and Linux. Intel OpenVINO, Qualcomm QNN and AMD Vitis AI on Windows | | Runs on | Windows, macOS on Apple silicon and Linux, x64 and ARM64. Node.js 20 or later, Python 3.11 to 3.14, .NET 8 or later, Rust 1.70 or later | | Models | A curated catalogue. The docs name GPT OSS, Qwen, DeepSeek, Mistral and Phi for chat and Whisper for transcription, plus embedding models. Each model has its own licence | | What leaves the machine | Model and execution provider downloads, and telemetry events through Microsoft's 1DS SDK, on by default. The privacy file says prompts, outputs and audio are not collected | | Releases in 90 days | 6 (CLI preview 0.10.2 on 14 July, v1.2.4, CLI preview 0.10.3, v2.0.1, v2.1.0, CLI preview 0.11.0 on 7 October 2026) | | Packages | foundry-local-sdk 2.1.0 on npm (105,372 downloads in the week to 4 October 2026), PyPI (47,344 in the last week) and crates.io, and Microsoft.AI.Foundry.Local 2.1.0 on NuGet. CLI through winget and Homebrew | | Issues | 76 open and 362 closed on 8 October 2026, with 32 open pull requests. No published security advisory and no CVE found at NVD | | Capabilities | inference.local, inference.open-weights, embed.text, speech.stt | | Tags | local, open-source, free, no-card, account-free, no-auth, openai-compatible, csharp, typescript, python, rust, npu, telemetry-on-by-default | | JSON | https://www.anchorterminal.com/api/v1/tools/foundry-local.json | ## Score breakdown (methodology v0.4, October 2026 research run) Assessed 2026-10-08 from public evidence against the published checklist (https://www.anchorterminal.com/benchmark/#checklist). Confidence: medium. Performance and Task success pending (no score, not in the total); the total is Σ(score × weight) ÷ 80 over the 7 assessed categories. "This run" is each category's share of the 100 points. | Category | Weight | This run | Score (0–100) | Points | | --- | --- | --- | --- | --- | | Reliability | 16% | 20 | 68 | 13.6 | | Performance | 10% | pending | pending | n/a | | Schema & documentation | 13% | 16.2 | 53 | 8.6 | | Agent ergonomics | 13% | 16.2 | 61 | 9.9 | | Security & auth | 14% | 17.5 | 39 | 6.8 | | Payments & pricing | 10% | 12.5 | 60 | 7.5 | | Task success | 10% | pending | pending | n/a | | Maintenance & community | 7% | 8.8 | 88 | 7.7 | | Transparency & trust (editorial 63, provenance 80) | 7% | 8.8 | 72 | 6.3 | | Negative events | up to −15 | up to −15 | none recorded | 0 | | **Total** | | | | **60.5 → C** | ### Why each score - Reliability 68: Read with the local-software lines, since Foundry Local runs in the owner's application or as a local daemon. We graded the SDK and its optional local server, and say where the preview CLI differs. Official packages on npm, PyPI, NuGet and crates.io, the CLI through winget and Homebrew, with Node.js 20, Python 3.11 to 3.14, .NET 8 and Rust 1.70 stated (20). The repository holds test suites for the C++ runtime and each SDK, but the only public GitHub Actions workflow is a build check for samples. Build and test run in Azure Pipelines whose results we couldn't see, and the pipeline file says its push trigger is switched off (12 of 25). 76 open issues against 362 closed. Open reports include a Linux ARM64 segmentation fault (#1182, 8 October), HTTP 500 on Qwen3.5 CUDA models (#1179), memory kept after unload (#1079) and several GPU and NPU detection failures from September with no reply (13 of 25). Versioned releases with notes on GitHub, and v2.0.1 has a breaking-changes section. There is no changelog file, and the CLI's REST reference warns of breaking changes without notice (10 of 15). SDK 2.1.0. The PyPI package still carries an Alpha classifier and the CLI is a public preview (13 of 15). - Performance: Pending. Latency is measured per call by our probes, which haven't run yet, so this run doesn't score it. Its weight is shared across the assessed categories until the first probe window closes. - Schema & documentation 53: Read with the API lines, for the local server and the SDKs. No OpenAPI file in the repository or the docs. The server follows OpenAI's shapes and the SDKs are typed (8 of 25). Microsoft Learn returns each page as Markdown when asked with `Accept: text/markdown`. No llms.txt on learn.microsoft.com or foundrylocal.ai (6 of 10). The overview says when to use the SDK and when the server, and states that Foundry Local is not meant for multi-user serving (14 of 20). The REST reference lists parameter types, optional flags and the allowed values of `ep` in prose (9 of 15). 41 samples across four languages and curl examples, with no documented error responses beyond a troubleshooting table (9 of 15). Dated release notes on GitHub. The REST reference, last updated on 14 July 2026, lists `/openai/*` and `/foundry/list` routes that the v2 source doesn't register and omits `/v1/responses`, and the Learn pages we read don't mention the Session API that 2.0.1 made the main interface and still give install commands for the `-winml` packages that release removed (7 of 15). - Agent ergonomics 61: Read with the API lines. Output is sized with `max_tokens` or `max_completion_tokens`, streaming is opt-in, and `/v1/responses` stores responses so a caller can send `previous_response_id` (18 of 25). `foundry model list` filters by device, task, search term and limit, and stored responses can be listed. Model lists over HTTP aren't paged (12 of 20). Errors in the source are OpenAI-shaped objects with a message and a type of `invalid_request_error` or `server_error`, with `code` always null and no documentation (9 of 20). Inference calls are stateless and safe to repeat. The README says cancellation is cooperative and that requests have no built-in timeout, and we found no retry guidance (8 of 20). An alias picks the best variant for the hardware, models download on first use, and SDKs exist in four documented languages. The server's port is dynamic by default, so a caller has to discover it (14 of 15). - Security & auth 39: Read with the tool checklist, for the local server. No credential exists. The server is off until the application or the CLI starts it, binds to 127.0.0.1 by default, and the docs tell clients to set auth to None (8 of 30). No read-only mode. Any local process that reaches the port can load and unload models and call `POST /shutdown`. Running in-process with no server is the default and removes the port altogether (6 of 20). The API returns model output, with no guidance on injected instructions (6 of 15). `foundry server logs` and SDK log levels give the operator a record, with no caller identity (7 of 15). SECURITY.md sends reports to MSRC and promises a reply within 24 hours, GitHub Actions are pinned to commit hashes, and no advisory or CVE was found. The security.txt on www.microsoft.com expired on 23 September 2026, and open issue #1123 says binaries downloaded at runtime rely on transport integrity alone (12 of 20). - Payments & pricing 60: Read with the self-hosted rule, since everything an agent calls runs on the owner's machine. No x402, MPP or L402 (0). The SDK is MIT, the CLI is a free download, and no account, card or Azure subscription is needed, so 20, 20 and 20 on the last three lines. - Task success: Pending. Task success needs the category task suites run through each tool, which haven't run yet, so this run doesn't score it. Its weight is shared across the assessed categories until then. A data provider's data-quality score is published on its listing now and becomes half of this category when it's scored. - Maintenance & community 88: SDK 2.1.0 reached the registries on 29 September 2026 and CLI preview 0.11.0 followed on 7 October (30). Six releases in the 90 days to 8 October (20). 132 commits on main since 10 July, the newest on 7 October, and 362 issues closed against 76 open. Many of the newest open issues had no comment when we read them, and SUPPORT.md is still Microsoft's unedited template (15 of 25). Current SDKs at the same version on npm, PyPI, NuGet and crates.io (15). Dependabot is configured, native dependency versions are pinned in one file that the build validates, and actions are pinned. We couldn't see CI results (8 of 10). - Transparency & trust 72: The editorial half. The SDKs and the v2 native runtime are MIT in a public repository. The CLI is closed under Microsoft Software Licence Terms, execution providers come under NVIDIA, Intel and Qualcomm licences, and each model has its own (24 of 30). The repository's privacy file says telemetry goes to Microsoft through the 1DS SDK and that prompts, outputs, audio, raw device identifiers and secrets are not collected. The Learn FAQ lists only downloads and optional diagnostics as network use and doesn't mention default telemetry, so the two statements don't fully agree, and no retention period is given (16 of 30). The v2.0.1 notes say the deprecated clients stay through 2026, and Learn keeps a legacy SDK reference and a migration guide. There is no written deprecation policy (9 of 20). Telemetry is disclosed in the README and is on by default. A config option or `ORT_TELEMETRY_DISABLED=1` turns off non-essential telemetry, and what counts as essential isn't stated (14 of 20). Fix list for a coding agent, everything this grade says the listing lacks, the biggest gain first (16 items): https://www.anchorterminal.com/fixes/foundry-local.md (JSON https://www.anchorterminal.com/fixes/foundry-local.json) ### What we couldn't check - Unchecked: whether the preview CLI's daemon adds the `/openai/*` and `/foundry/list` routes of the REST reference on top of the v2 runtime's routes. The CLI is closed source and we didn't run it - Unchecked: CI results. Build and test run in Azure Pipelines, which we couldn't read, and api.github.com refused our requests for rate limiting - Unchecked: the Microsoft Privacy Statement, which answered 403, so retention and sharing terms for the telemetry were not read - Unchecked: how many models the catalogue holds. foundrylocal.ai/models is drawn by script - What the telemetry events contain and which of them count as essential. The privacy file doesn't list them - Whether the local server checks the Origin or Host header. We found no such check in the service source but did not test it - The star count (2.6k) and issue counts come from GitHub's web pages as read through a summarising fetcher ### Sources - overview and FAQ: (seen 2026-10-08) - architecture: (seen 2026-10-08) - REST reference (CLI preview): (seen 2026-10-08) - CLI reference: (seen 2026-10-08) - SDK reference: (seen 2026-10-08) - REST server how-to: (seen 2026-10-08) - best practice and troubleshooting: (seen 2026-10-08) - get started: (seen 2026-10-08) - repository (README, LICENSE, SECURITY.md, SUPPORT.md, workflows, tags), cloned at commit of 7 October 2026: (seen 2026-10-08) - v2 web service source (routes, binding): (seen 2026-10-08) - privacy file: (seen 2026-10-08) - licence file: (seen 2026-10-08) - releases: (seen 2026-10-08) - v2.0.1 release notes: (seen 2026-10-08) - v2.1.0 release notes: (seen 2026-10-08) - open issues: (seen 2026-10-08) - security advisories (none published): (seen 2026-10-08) - npm latest: (seen 2026-10-08) - npm weekly downloads: (seen 2026-10-08) - PyPI release history: (seen 2026-10-08) - PyPI downloads: (seen 2026-10-08) - crates.io: (seen 2026-10-08) - NuGet versions: (seen 2026-10-08) - security.txt (expired 23 September 2026): (seen 2026-10-08) - product site: (seen 2026-10-08) - NVD keyword search (no Foundry Local CVE): (seen 2026-10-08) - RDAP for microsoft.com: (seen 2026-10-08) ## Who's behind it (provenance 80/100, checked 2026-10-08) | Check | Finding | Points | | --- | --- | --- | | Legal entity named | Microsoft Corporation | 20/20 | | Domain age | microsoft.com, registered 1991-05-02 (35 years) | 15/15 | | Endpoint on the vendor's domain | no hosted endpoint | n/a | | Terms of service | nothing hosted, so the MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own licence stands in | 10/10 | | Privacy policy | nothing hosted, not scored | n/a | | Status page | not found | 0/10 | | Changelog | published | 10/10 | | security.txt | published but past its Expires date | 5/10 | The repository is under GitHub's microsoft organisation, the licence file reads Copyright (c) Microsoft Corporation, and foundrylocal.ai names Microsoft Corporation as publisher. The terms link is the repository's LICENSE file, which holds the MIT licence for the SDK and the Microsoft Software Licence Terms for the CLI. Microsoft publishes no other agreement for Foundry Local that we found. The privacy link is the product's own privacy file in the repository, which the README links. It and the CLI licence both refer on to the Microsoft Privacy Statement, a company-wide notice, which answered 403 to our request. www.microsoft.com/.well-known/security.txt loads and points to the MSRC researcher portal, but its Expires field is 2026-09-23T16:00:00.000Z, which had passed on 8 October 2026. learn.microsoft.com and foundrylocal.ai return 404. No status page is listed because the software runs on the owner's machine. The model catalogue is a cloud service with no status page that we found. RDAP gives microsoft.com a registration date of 1991-05-02 and foundrylocal.ai one of 2025-10-07. ### Terms and privacy, as read A reading by a fixed set of rules, each answered with the vendor's own sentence. Not legal advice. **Terms of service**. Nothing is hosted by the vendor, so there are no terms of service to read. The MIT for the SDKs and the v2 native runtime. The CLI is closed source under Microsoft Software Licence Terms. Execution providers carry NVIDIA, Intel and Qualcomm licences, and each model carries its own licence stands in and the check scores in full. **Privacy policy**. Nothing is hosted by the vendor, so there is no privacy policy to read and the check isn't scored. ## Probe metrics Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. Live uptime, where we poll the endpoint, is under Live and doesn't change the score. ## Strengths - SDK 2.1.0 published on npm, PyPI, NuGet and crates.io on 29 September 2026, with the SDK and the v2 native runtime under MIT in a public repository - A model alias selects the best variant for the machine's CPU, GPU or NPU, with CUDA, WebGPU, OpenVINO, QNN and Vitis AI execution providers - The v2 server answers `/v1/chat/completions`, `/v1/responses`, `/v1/embeddings`, `/v1/audio/transcriptions` and `/v1/models` in OpenAI's shapes - Six releases in the 90 days to 8 October 2026, and the v2.0.1 notes carry a breaking-changes section - No account, key, card or Azure subscription is needed, and prompts and outputs are processed on the device per the docs ## Weaknesses - The local server takes no credential, and its routes include model load and unload and `POST /shutdown` - The REST reference on Microsoft Learn lists `/openai/*` and `/foundry/list` routes that the v2 runtime source doesn't register, and doesn't cover `/v1/responses` - Telemetry is on by default through Microsoft's 1DS SDK. The opt-out covers non-essential telemetry only, and the Learn FAQ doesn't mention it - The CLI is a closed-source public preview, and its REST reference warns of breaking changes without notice - 76 open issues on 8 October 2026, among them a Linux ARM64 segmentation fault (#1182) and several unanswered GPU and NPU detection reports ## Before you call it (notes for agents) 1. Read the server URL from `manager.urls[0]` or `foundry server status`. The port is dynamic unless the owner sets `web.urls` or `foundry server start --port` 2. Send the model ID that `GET /v1/models` returns, not the alias. The alias resolves to a hardware-specific variant 3. Check `supportsToolCalling` before sending tools. Support differs by variant, and issue #1183 reports Qwen tool calling failing on QNN 4. Set your own request timeout. Inference has no built-in one, and cancellation takes effect only after the current generation step 5. Ask the owner to set `ORT_TELEMETRY_DISABLED=1` or `disableNonessentialTelemetry` before the manager is created if telemetry must be off ## Connect Install: ```bash pip install foundry-local-sdk # or: npm install foundry-local-sdk # CLI (preview): winget install Microsoft.FoundryLocal # macOS: brew tap microsoft/foundrylocal && brew install foundrylocal ``` First request: ```bash foundry server start --port 39839 --idle-timeout 0 curl http://localhost:39839/v1/chat/completions \ -H "Content-Type: application/json" \ -d '{"model": "", "messages": [{"role": "user", "content": "What is the golden ratio?"}]}' ``` Through letme (picks today, calling later): https://letme.dev/foundry-local. letme answers with the pick and how to call it direct; calling through letme (one key, the vendor's own price) comes later. How it works: https://www.anchorterminal.com/letme/index.md ## Similar tools Ranked by shared capabilities, then score. Same-category tools with no shared capability key are listed last. | Tool | Grade | Score | Rank | Shared capabilities | x402 | Markdown | | --- | --- | --- | --- | --- | --- | --- | | LocalAI | B | 68 | 216 | inference.local, inference.open-weights, embed.text, speech.stt | no | https://www.anchorterminal.com/tools/localai.md | | Lemonade | B | 63.8 | 336 | inference.local, inference.open-weights, embed.text, speech.stt | no | https://www.anchorterminal.com/tools/lemonade.md | | KoboldCpp | C | 60.5 | 462 | inference.local, inference.open-weights, embed.text, speech.stt | no | https://www.anchorterminal.com/tools/koboldcpp.md | | DeepInfra | B | 63 | 371 | inference.open-weights, embed.text, speech.stt | no | https://www.anchorterminal.com/tools/deepinfra.md | | llama.cpp | C | 60.2 | 476 | inference.local, inference.open-weights, embed.text | no | https://www.anchorterminal.com/tools/llama-cpp.md | | LM Studio | C | 57.8 | 536 | inference.local, inference.open-weights, embed.text | no | https://www.anchorterminal.com/tools/lm-studio.md | ## Panel reviews (0) Reviewed by the Anchor panel (https://www.anchorterminal.com/reviewers/index.md): . Desk reviews, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. For a desk review, the outcome says whether the reviewer's questions could be answered from public material: success, partial or failure. How reviews work: https://www.anchorterminal.com/reviews/how-it-works.md ## Notable - Version 2.0.1 (GitHub release dated 1 September 2026) replaced the in-process OpenAI-style clients with `ChatSession`, `EmbeddingsSession` and `AudioSession`, merged the `-winml` packages into one per language, and keeps the old clients through 2026 (source: ) - The v2 runtime registers `/v1/chat/completions`, `/v1/responses`, `/v1/embeddings`, `/v1/audio/transcriptions`, `/v1/models`, `/models/load/{name}`, `/models/unload/{name}`, `/models/loaded`, `/status` and `POST /shutdown`, with no credential check, and binds to 127.0.0.1 unless another host is configured (source: ) - The REST reference (page updated 14 July 2026) describes `/openai/status`, `/foundry/list`, `/openai/download` and `/openai/load/{name}`, none of which appear in the repository source on 8 October 2026 (source: ) - Telemetry is enabled by default through the 1DS SDK. The privacy file says prompts, model outputs, audio contents, raw device identifiers and secrets are not collected, and `ORT_TELEMETRY_DISABLED=1` turns off non-essential telemetry (source: ) - Microsoft's FAQ says Foundry Local targets one user on one device and is not a server inference stack, with no request queuing or batching for concurrent clients (source: ) - The SDK is MIT. The CLI is under Microsoft Software Licence Terms that forbid reverse engineering and sharing, and cap damages at US $5 (source: ) - CLI preview 0.11.0 of 7 October 2026 moves the CLI to SDK 2.1.0, ONNX Runtime 1.30.0 and ONNX Runtime GenAI 0.17.1 (source: ) ## Compare - [AnythingLLM vs Foundry Local](https://www.anchorterminal.com/compare/anythingllm-vs-foundry-local.md): D 53.3 vs C 60.5 - [Docker Model Runner vs Foundry Local](https://www.anchorterminal.com/compare/docker-model-runner-vs-foundry-local.md): C 57.1 vs C 60.5 - [Foundry Local vs Core](https://www.anchorterminal.com/compare/foundry-local-vs-ghost-core.md): C 60.5 vs F 7.3 - [Foundry Local vs GPT4All](https://www.anchorterminal.com/compare/foundry-local-vs-gpt4all.md): C 60.5 vs F 36.2 - [Foundry Local vs Jan](https://www.anchorterminal.com/compare/foundry-local-vs-jan.md): C 60.5 vs D 51.3 - [Foundry Local vs Khoj](https://www.anchorterminal.com/compare/foundry-local-vs-khoj.md): C 60.5 vs E 38.5 - [Foundry Local vs KoboldCpp](https://www.anchorterminal.com/compare/foundry-local-vs-koboldcpp.md): C 60.5 vs C 60.5 - [Foundry Local vs Lemonade](https://www.anchorterminal.com/compare/foundry-local-vs-lemonade.md): C 60.5 vs B 63.8 - [Foundry Local vs llama.cpp](https://www.anchorterminal.com/compare/foundry-local-vs-llama-cpp.md): C 60.5 vs C 60.2 - [Foundry Local vs LM Studio](https://www.anchorterminal.com/compare/foundry-local-vs-lm-studio.md): C 60.5 vs C 57.8 - [Foundry Local vs LocalAI](https://www.anchorterminal.com/compare/foundry-local-vs-localai.md): C 60.5 vs B 68 - [Foundry Local vs MLX LM](https://www.anchorterminal.com/compare/foundry-local-vs-mlx-lm.md): C 60.5 vs D 52.2 - [Foundry Local vs Ollama](https://www.anchorterminal.com/compare/foundry-local-vs-ollama.md): C 60.5 vs C 56.3 - [Foundry Local vs Open WebUI](https://www.anchorterminal.com/compare/foundry-local-vs-open-webui.md): C 60.5 vs D 51.8 - [Foundry Local vs screenpipe](https://www.anchorterminal.com/compare/foundry-local-vs-screenpipe.md): C 60.5 vs C 60.8 - [Foundry Local vs TextGen](https://www.anchorterminal.com/compare/foundry-local-vs-text-generation-webui.md): C 60.5 vs E 45.1 - [Foundry Local vs Underdog](https://www.anchorterminal.com/compare/foundry-local-vs-underdog.md): C 60.5 vs F 29.5 ## Verify this listing For the vendor. The badge or a plain link to this page verifies the listing, from a page on microsoft.com or foundrylocal.ai or one of their subdomains, or the README of github.com/microsoft/Foundry-Local. It shows the listing is the vendor's and that the vendor knows it's here, and it never changes a grade, rank or review. The vendor sends the page's address to `POST https://www.anchorterminal.com/api/v1/verify` as `{"slug": "foundry-local", "url": "…"}`, or calls the `verify_listing` tool at https://www.anchorterminal.com/mcp. We fetch the page once, then again every week; two failed checks in a row and the verification lapses, and a later pass restores it. What we check: https://www.anchorterminal.com/builders/index.md#verify HTML badge: ```html Foundry Local on Anchor Terminal ``` Markdown badge, for a README: ```markdown [![Foundry Local on Anchor Terminal](https://www.anchorterminal.com/badges/foundry-local.svg)](https://www.anchorterminal.com/tools/foundry-local) ``` Plain link: ```html Foundry Local on Anchor Terminal ``` ## Share this listing For the vendor. Sharing assets for social media, two PNGs of 1200 × 630 that say Foundry Local is listed on Anchor Terminal, with the vendor's logo and this page's address and no grade or score. - Dark: https://www.anchorterminal.com/assets/share/foundry-local-dark.png - Light: https://www.anchorterminal.com/assets/share/foundry-local-light.png