Category · Agent frameworks
Agent frameworks and SDKs
Libraries that run the agent loop. Tool calling, MCP, multi-agent hand-offs, durable state, human approval and tracing. Compared on what they do by default, including the telemetry they send.
Capability keys agent.framework · agent.multi-agent · agent.durable · agent.mcp-client · All tools
letme.dev/agent.framework picks the top-graded tool in this list and says how to call it direct; calling through letme comes later.
The same listing from the live API. Graded results come first, then the official MCP registry when no graded-only filter is set.
https://www.anchorterminal.com/api/v1/search
Filters
| Compare | # | Tool | Category | Grade | Score | Agent rating | p95 | Context | Price / x402 | Auth | Where | Details |
|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | OpenAI Agents SDKOpenAI · Agent framework | Frameworks | AA | 86.5 | 3.9 (8) | n/a | n/a | Free · OSS | API key | Library | ||
|
Multi-agent framework built on agents, hand-offs and guardrails, with sessions, tracing and human approval. Top strength MCP in about 11 lines, with static and dynamic tool filters and require_approval Top weakness Tracing on by default, with model and tool content, sent to OpenAI |
||||||||||||
| 7 | Pydantic AIPydantic · Agent framework | Frameworks | A | 80 | 4.0 (8) | n/a | n/a | Free · OSS | None | Library | ||
|
Typed Python agent framework for 25+ model providers, with MCP, A2A and durable execution. Top strength Typed outputs and tools, validated by Pydantic, with failed validations sent back to the model Top weakness Python only |
||||||||||||
| 45 | Agent Development Kit (ADK)Google · Agent framework | Frameworks | BB | 74.9 | 3.0 (8) | n/a | n/a | Free · OSS | None | Library | ||
|
Google's code-first toolkit to build, evaluate and deploy agents in Python, TypeScript, Go, Java and Kotlin. Top strength Python, TypeScript, Go, Java and Kotlin Top weakness Two critical CVEs in 2026, one of them in tool confirmation itself |
||||||||||||
| 71 | Claude Agent SDKAnthropic · Agent framework | Frameworks | BB | 72.4 | none | n/a | n/a | Free | API key | Library | ||
|
Runs the Claude Code agent as a library, with its loop, built-in tools, permissions, sessions, hooks and subagents. Top strength Claude Code's file, shell, search and web tools, with six permission modes, a canUseTool callback and PreToolUse hooks Top weakness Claude models only, so a person has to set up a Claude API key or cloud account first |
||||||||||||
| 95 | LangGraphLangChain · Agent framework | Frameworks | BB | 70.6 | 3.5 (2) | n/a | n/a | Free · OSS | None | Library | ||
|
Low-level runtime for long-running, stateful agent graphs with checkpointed persistence, human-in-the-loop, memory and streaming. Top strength Checkpointed persistence, so runs survive restarts and resume where they stopped Top weakness No MCP client of its own. LangChain's langchain.mcp is in beta |
||||||||||||
| 149 | CrewAICrewAI · Agent framework | Frameworks | B | 67 | 3.0 (2) | n/a | n/a | Free · OSS | None | Library | ||
|
Python framework for teams of role-playing agents in Crews, plus event-driven Flows for stateful workflows. Top strength An MCP server in five lines through the mcps field, with static and dynamic tool filters Top weakness Anonymous telemetry on by default, with no stated destination or retention |
||||||||||||
Nothing matches these filters. .
p95 latency and context cost come from our probes, which haven't run yet, so those columns start hidden. Grades run from AA to F, and agent-ready means BB or better. Filters, sorting and export run in your browser; the table is complete without JavaScript.
Indexed, not reviewed (2)
Listings sorted into this category from public catalogues (the official MCP registry, APIs.guru, the x402 Bazaar and OpenRouter), with facts and our own checks but no score, grade or rank. How the index works.
| Listing | Kind | What it does | Why it's here |
|---|---|---|---|
| FlurryPORT flurryport.io | MCP server | Webhook capture and replay, delivery pipes, and signed multi-agent rooms. No signup to start. | vendor's own |
| Gemot Deliberation Server gemot.dev | MCP server | Deliberation primitive for multi-agent coordination: cruxes, vote clustering, consensus. | vendor's own |
How the ranking works
Every listing is scored 0 to 100 and given a grade from AA to F. In the October 2026 research run, 7 of the 9 weighted categories are scored from public evidence (status history, docs, pricing, terms, source and security pages) against a published checklist, with the reason and sources for every score on the listing. Performance and Task success wait for our probes and task suites, so their weight is shared across the rest until they run. Negative events deduct up to 15 points. Read the methodology.
For companies
Do agents find, use and choose your tools?
An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.
