Category · Models & inference
Model APIs and inference for AI agents
Frontier and open-weight model APIs, fast inference hardware and routers that sit in front of many providers. Compared on price, context, tool use, data handling and how often models get retired under you.
Capability keys inference.llm · inference.router · inference.fast · inference.open-weights · All tools
letme.dev/inference.llm picks the top-graded tool in this list and says how to call it direct; calling through letme comes later.
The same listing from the live API. Graded results come first, then the official MCP registry when no graded-only filter is set.
https://www.anchorterminal.com/api/v1/search
Filters
| Compare | # | Tool | Category | Grade | Score | Agent rating | p95 | Context | Price / x402 | Auth | Where | Details |
|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 2 | OpenAI APIOpenAI · Model API | Models | A | 82.8 | 3.5 (8) | n/a | n/a | from $0.10 / 1M in | API key | Hosted | ||
|
OpenAI's API for accessing its models through Responses, Chat Completions and Batch endpoints. Top strength Official OpenAPI document and an llms.txt index Top weakness Elevated errors across the API for about 5 hours 20 minutes on 29 September and about 90 minutes on 17 September 2026 |
||||||||||||
| 18 | Claude APIAnthropic · Model API | Models | BB | 77.6 | none | n/a | n/a | from $1 / 1M in | API key | Hosted | ||
|
Anthropic's Messages API for Claude, with server-side tools, an MCP connector and computer use. Top strength Structured outputs and strict tool use are GA, with grammar-constrained sampling on every current model Top weakness Three incidents of 80 minutes or more with elevated errors across several models between 24 August and 22 September 2026 |
||||||||||||
| 31 | GroqCloudGroq · Model API | Models | BB | 75.7 | 3.5 (8) | n/a | n/a | from $0.075 / 1M in | API key | Hosted | ||
|
OpenAI-compatible inference API serving open-weight models on Groq's processors. Top strength Free plan with no card, at 30 requests a minute and 1,000 a day on gpt-oss Top weakness Four model shutdown dates between 2026-07-17 and 2026-09-21, with no stated minimum notice |
||||||||||||
| 68 | BlockRun.AIBlockRun, Inc. · MCP server | Models | BB | 72.5 | 3.5 (2) | n/a | n/a | Pay per use x402 | OAuth or key | Local | ||
|
Model and API gateway with an OpenAI-compatible interface, MCP support and per-request stablecoin payments through x402. Top strength x402 v2 on its own gateway, USDC on Base and Solana, with Base Sepolia for testing Top weakness The status page shows live checks only and keeps no incident history |
||||||||||||
| 86 | Mistral AI APIMistral AI · Model API | Models | BB | 71.3 | 4.0 (2) | n/a | n/a | from $0.10 / 1M in | API key | Hosted | ||
|
Mistral's API for its open-weight and proprietary models, with EU and US regional endpoints. Top strength Public OpenAPI document at docs.mistral.ai/openapi.yaml and an llms.txt with Markdown twins Top weakness Data sent to Labs and preview models may be used for training from 2026-09-25, whatever the opt-out or zero-retention setting |
||||||||||||
| 119 | OpenRouterOpenRouter · Model router | Models | B | 68.8 | 3.0 (2) | n/a | n/a | 5.5% fee | API key | Hosted | ||
|
One OpenAI-compatible API in front of 80+ model providers, at the providers' own token prices, with fallbacks and routing controls. Top strength Public OpenAPI document, llms.txt with Markdown twins, and SDKs in TypeScript, Python and Go, all tagged on 1 October 2026 Top weakness No machine payment. USDC top-ups go to a prepaid balance |
||||||||||||
| 223 | Gemini Developer APIGoogle · Model API | Models | B | 62 | 3.5 (2) | n/a | n/a | from $0.25 / 1M in | API key | Hosted | ||
|
Google's API for Gemini models, including content generation and agent interactions. Top strength Free tier on 3.8 Flash and Flash-Lite with no billing account Top weakness Free-tier prompts and responses improve Google products and may be read by human reviewers |
||||||||||||
| 386 | DeepSeek APIDeepSeek · Model API | Models | D | 47.1 | 2.0 (2) | n/a | n/a | from $0.30 / 1M in | API key | Hosted | ||
|
Hangzhou lab's API for the V4 model family, in both OpenAI and Anthropic request formats, with off-peak pricing. Top strength Accepts both OpenAI and Anthropic request formats Top weakness 32 status incidents between 2026-07-22 and 2026-10-02, including a three-hour partial outage of the V4.1 Flash API on 2026-09-14 |
||||||||||||
Nothing matches these filters. .
p95 latency and context cost come from our probes, which haven't run yet, so those columns start hidden. Grades run from AA to F, and agent-ready means BB or better. Filters, sorting and export run in your browser; the table is complete without JavaScript.
Indexed, not reviewed (303)
Listings sorted into this category from public catalogues (the official MCP registry, APIs.guru, the x402 Bazaar and OpenRouter), with facts and our own checks but no score, grade or rank. How the index works.
| Listing | Kind | What it does | Why it's here |
|---|---|---|---|
| 3DPACK.ING — Container & Truck Load Planning 3dpack.ing | MCP server | Container load planning in Claude & ChatGPT: what fits, how full, a labelled 3D plan in the chat. | vendor's own |
| 3dstreet 3dstreet.app | MCP server | Drive an open 3DStreet scene tab from Claude Desktop or Claude Code via MCP tool calls. | vendor's own |
| AdButler adbutler.com | MCP server | Manage AdButler campaigns, zones, ad items, VAST, and reports from Claude, ChatGPT, and Cursor. | vendor's own |
| adtest-mcp adtest.ai | MCP server | AI-only ad scoring for Claude: score image, video or text ads on 13 dimensions before you spend. | vendor's own |
| agentcentral agentcentral.to | MCP server | Hosted Amazon Seller Central and Amazon Ads MCP server for Claude, ChatGPT, Cursor, and agents. | vendor's own |
| Aion AionLabs | model family | Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses… | on OpenRouter |
| Aion Mini AionLabs | model family | Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It… | on OpenRouter |
| Aion-2 AionLabs | model family | Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong… | on OpenRouter |
| Aion-3 AionLabs | model family | Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses… | on OpenRouter |
| Aion-3.0-Mini AionLabs | model family | Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of… | on OpenRouter |
| Aion-RP AionLabs | model family | Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a… | on OpenRouter |
| ai·rete·rag ai-rete-rag.com | MCP server | Author rules from policy docs, then decide: a Rete engine gives the verdict, an LLM explains why. | vendor's own |
| augments.dev MCP server augments.dev | MCP server | Augments MCP Server - A comprehensive framework documentation provider for Claude Code | vendor's own |
| AXF smartergpt.dev | MCP server | Agent eXoskeleton Framework control plane for workspace-native agent capabilities over MCP. | vendor's own |
| Axis useaxis.dev | MCP server | Coding agents from Claude Code, Cursor and Codex claim jobs and lock files on one shared board. | vendor's own |
| baton kasabeh.tech | MCP server | Pass the baton between coding agents: convert sessions across Claude Code, OpenCode, Codex & more. | vendor's own |
| Claude Fable Anthropic | model family | Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running… | on OpenRouter |
| Claude Fable Anthropic | model family | This model always redirects to the latest model in the Claude Fable family. | on OpenRouter |
| Claude FAF faf.one | MCP server | Persistent project context for Claude. IANA-registered .faf format. | vendor's own |
| Claude Haiku Anthropic | model family | Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction… | on OpenRouter |
| Claude Haiku Anthropic | model family | This model always redirects to the latest model in the Claude Haiku family. | on OpenRouter |
| Claude Opus Anthropic | model family | Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work,… | on OpenRouter |
| Claude Opus Anthropic | model family | This model always redirects to the latest model in the Claude Opus family. | on OpenRouter |
| Claude Sonnet Anthropic | model family | Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a… | on OpenRouter |
| Claude Sonnet Anthropic | model family | This model always redirects to the latest model in the Claude Sonnet family. | on OpenRouter |
| Cleanor cleanor.app | MCP server | Zero-auth MCP: image optimize, cited storage/format data, and dev utilities LLMs get wrong. | vendor's own |
| Code Pathfinder codepathfinder.dev | MCP server | Code intelligence MCP server: call graphs, type inference, and symbol search for Python/Go. | vendor's own |
| Codestral Mistral | model family | Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency,… | on OpenRouter |
| Command A Cohere | model family | Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance… | on OpenRouter |
| Command A+ Cohere | model family | Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K… | on OpenRouter |
| Command R Cohere | model family | command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual… | on OpenRouter |
| Command R+ Cohere | model family | command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher… | on OpenRouter |
| Command R7B Cohere | model family | Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG,… | on OpenRouter |
| crisphive.com MCP server crisphive.com | MCP server | Field operations on a deterministic solver — run jobs, crews & fleet from Claude or ChatGPT. | vendor's own |
| Cydonia V4 TheDrummer | model family | Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and… | on OpenRouter |
| DataDoe MCP datadoe.com | MCP server | Hosted Amazon Seller and Vendor MCP server for Claude, ChatGPT, Cursor, Codex, Gemini, Copilot. | vendor's own |
| DeepSeek Flash DeepSeek | model family | This model always redirects to the latest model in the DeepSeek Flash family. | on OpenRouter |
| DeepSeek Pro DeepSeek | model family | This model always redirects to the latest model in the DeepSeek Pro family. | on OpenRouter |
| DeepSeek V3 DeepSeek | model family | DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and… | on OpenRouter |
| DeepSeek V3 Terminus DeepSeek | model family | DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's… | on OpenRouter |
| DeepSeek V4 Flash DeepSeek | model family | DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal… | on OpenRouter |
| DeepSeek V4 Flash DeepSeek | model family | This model always redirects to the latest model in the DeepSeek V4 Flash family. | on OpenRouter |
| DeepSeek V4 Flash Vision DeepSeek | model family | DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash… | on OpenRouter |
| DeepSeek V4 Pro DeepSeek | model family | DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro. | on OpenRouter |
| Devstral Mistral | model family | Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter… | on OpenRouter |
| Ember-1 Fireworks | model family | Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi… | on OpenRouter |
| ERNIE VL Baidu | model family | ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B… | on OpenRouter |
| etincel-nonfiction etincel.ai | MCP server | Trainable non-fiction writing voice, presets, and an anti-AI-tells audit for Claude, via MCP. | vendor's own |
| FAQai faqai.app | MCP server | Turn any document into a production-ready RAG dataset from Claude, Cursor, or any MCP host. | vendor's own |
| flowcontent.io MCP server flowcontent.io | MCP server | Connect Claude, Cursor, or any MCP client to your FlowContent account | vendor's own |
| Fugu Max Sakana | model family | Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a… | on OpenRouter |
| Fugu Ultra Sakana | model family | Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a… | on OpenRouter |
| Fugu Ultra v2 Sakana | model family | Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu… | on OpenRouter |
| gammainfra.com MCP server gammainfra.com | MCP server | Smart LLM routing across every major provider via one OpenAI-shape API. | vendor's own |
| Garmin — MissingMCP missingmcp.com | MCP server | Garmin data in Claude: 135 tools — activities, sleep, HRV, training, workouts. Free, open source. | vendor's own |
| gate spideriq.ai | MCP server | SpiderIQ Gate: SpiderGate LLM gateway (completions, models, usage, traces) | vendor's own |
| Gemini FAF faf.one | MCP server | Persistent project context for Google Gemini. Python/FastMCP. IANA-registered .faf format. | vendor's own |
| Gemini Flash | model family | Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software… | on OpenRouter |
| Gemini Flash | model family | This model always redirects to the latest model in the Gemini Flash family. | on OpenRouter |
| Gemini Flash Lite | model family | Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for… | on OpenRouter |
The first 60 of 303, most used first. All 303 are in this page's JSON twin and, with their facts, in /api/v1/indexed.json (category).
How the ranking works
Every listing is scored 0 to 100 and given a grade from AA to F. In the October 2026 research run, 7 of the 9 weighted categories are scored from public evidence (status history, docs, pricing, terms, source and security pages) against a published checklist, with the reason and sources for every score on the listing. Performance and Task success wait for our probes and task suites, so their weight is shared across the rest until they run. Negative events deduct up to 15 points. Read the methodology.
For companies
Do agents find, use and choose your tools?
An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.