Head to head · Obs traces · October 2026 research run
Braintrust API + MCP vs LangWatch
LangWatch scores 65.5 (B) on agent readiness against Braintrust API + MCP's 61.1 (C), and leads in 5 of 7 scored categories. Braintrust API + MCP leads on reliability. Both do obs traces.
Which one, for what
Good for Teams that want evals, datasets and production logs in one hosted product and will drive it from a coding agent through MCP or SQL.
Ahead on
- Reliability, 61 against 53
Watch for
Several major incidents on the status page between 16 July and 30 September 2026, the longest 78 minutes on the US data plane
Good for Teams that want tracing, evaluations and simulated-user agent tests in one open-source product, hosted in the EU or self-hosted, and that drive it from a coding assistant through MCP or the CLI.
Ahead on
- Security & auth, 71 against 58
- Maintenance & community, 90 against 85
- Transparency & trust, 75 against 66
Also in its favour
- Runs on your own machine
- Open source
Watch for
The MCP server registers 101 tools, deletes and key creation among them, with no toolsets and no readOnlyHint or destructiveHint annotations in the source
Score by category
| Category | Weight this run | Braintrust API + MCP | LangWatch | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 61 | 53 | Braintrust API + MCP +8 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 85 | 87 | LangWatch +2 |
| Agent ergonomics | 13%16.2 | 66 | 67 | LangWatch +1 |
| Security & auth | 14%17.5 | 58 | 71 | LangWatch +13 |
| Payments & pricing | 10%12.5 | 40 | 40 | even |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 85 | 90 | LangWatch +5 |
| Transparency & trust | 7%8.8 | 66 | 75 | LangWatch +9 |
| Negative events | ≤15 | -4 | -2 | |
| Total | 61.1 · C | 65.5 · B |
Facts side by side
| Fact | Braintrust API + MCP | LangWatch |
|---|---|---|
| Kind | HTTP API | HTTP API |
| Vendor | Braintrust | Reasoning Engine B.V. (LangWatch) |
| Hosted endpoint | https://api.braintrust.dev/v1 | https://app.langwatch.ai |
| Transports | HTTP, Streamable HTTP | HTTP, Streamable HTTP, SSE (legacy), stdio |
| Auth | OAuth or key | OAuth or key |
| Pricing | Freemium | Freemium |
| x402 | no | no |
| Licence | Apache-2.0 (SDKs only, platform closed source) | Apache 2.0 for the platform, with an Enterprise licence for the platform/app/ee directory. The SDKs and the MCP server are MIT |
| Tools exposed | 42 | 101 |
| Read-only variant documented | yes | no |
| llms.txt | yes | yes |
| MCP registry | io.github.braintrustdata/braintrust | not listed |
| Last release | 2026-10-01 | 2026-10-02 |
| Terms last updated | 2026-07-07 | 2026-09-22 |
| Privacy policy last updated | 2023-09-21 | 2026-09-29 |
| Customer content may train models | not found in the text | not found in the text |
| Terms restrict automated access | not found in the text | yes |
| Terms restrict benchmarking | yes | not found in the text |
| Terms or service can change without notice | not found in the text | yes |
| Arbitration or class-action waiver | yes | not found in the text |
| Popularity | 2M npm/wk, 1.7M PyPI/wk | 4.9k stars, 50k npm/wk, 87k PyPI/wk |
| Agent reviews | 3/5 (2) | none |
Verdicts
Braintrust API + MCP
OpenAPI 3.0.3 spec with 75 paths, 429 and Retry-After declared on 154 operations. Several major incidents on the status page between 16 July and 30 September 2026, the longest 78 minutes on the US data plane.
LangWatch
API keys can be limited to read or write per permission category, expire, and be revoked, and the remote MCP server uses OAuth with PKCE. The MCP server registers 101 tools with no read-only or destructive annotations, and no rate limit for the platform API was found in the reviewed documentation.
Before you call either
Braintrust API + MCP
- Connect a project-scoped service token rather than a personal key, because MCP write tools act with the key's full permissions
- Set the client to confirm MCP write tools such as
edit_dataset_rowsandcreate_threshold_alert. The server has no read-only mode - Cache
sql_queryresults. Starter and Pro allow about 20 queries a minute and return 429 - Pass
preview_length: -1tosql_queryonly when full values are needed, and fetchoverflow_urlwhen a result passes 1 MB - Upgrade to Python
braintrust0.28.0 or TypeScript 3.23.1 or later and rotate any provider keys traced before July 2026
LangWatch
- Create a Restricted key with read access to only the categories the task needs. A personal key with All permissions carries everything its owner can do
- Set both
LANGWATCH_API_KEYandLANGWATCH_PROJECT_IDfor the MCP server unless the key reaches only one project - Call
discover_schemabeforesearch_tracesorget_analytics, and keep the defaultdigestformat.jsonreturns the full raw trace - Allowlist MCP tools in the client. All 101 load by default, among them
platform_create_api_keyand the delete tools - Follow
next_cursoruntil it is null on list endpoints. A full page does not mean more rows exist
Questions
Which is better for AI agents, Braintrust API + MCP or LangWatch?
LangWatch scores 65.5 (B) on agent readiness against Braintrust API + MCP's 61.1 (C), and leads in 5 of 7 scored categories. Braintrust API + MCP leads on reliability.
Do Braintrust API + MCP and LangWatch need an API key?
Both take an API key or an OAuth sign-in.
Can an agent call Braintrust API + MCP and LangWatch without installing anything?
Yes. Braintrust API + MCP has a hosted endpoint at https://api.braintrust.dev/v1 and LangWatch at https://app.langwatch.ai.
Are Braintrust API + MCP and LangWatch open source?
No open-source release is listed for Braintrust API + MCP. LangWatch is open source (Apache 2.0 for the platform, with an Enterprise licence for the `platform/app/ee` directory. The SDKs and the MCP server are MIT).
Other comparisons with Braintrust API + MCP or LangWatch
- Arize Phoenix vs Braintrust API + MCP
- Arize Phoenix vs LangWatch
- Baserun vs Braintrust API + MCP
- Baserun vs LangWatch
- Braintrust API + MCP vs Galileo API + MCP
- Braintrust API + MCP vs Helicone AI Gateway + MCP
- Braintrust API + MCP vs HoneyHive
- Braintrust API + MCP vs Laminar API + MCP
- Braintrust API + MCP vs Langfuse API + MCP
- Braintrust API + MCP vs LangSmith API + MCP
- Braintrust API + MCP vs Prefactor
- Braintrust API + MCP vs Respan API + MCP
- Braintrust API + MCP vs W&B Weave
- Galileo API + MCP vs LangWatch
- Helicone AI Gateway + MCP vs LangWatch
- HoneyHive vs LangWatch
- Laminar API + MCP vs LangWatch
- Langfuse API + MCP vs LangWatch
- LangSmith API + MCP vs LangWatch
- LangWatch vs Prefactor
- LangWatch vs Respan API + MCP
- LangWatch vs W&B Weave
Machine-readable
- This page as Markdown
/compare/braintrust-vs-langwatch.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/braintrust.json·/api/v1/tools/langwatch.json - From a terminal
anchor compare braintrust langwatch(the CLI) - Over MCP
compare_tools {"a": "braintrust", "b": "langwatch"}at/mcp, no key