Head to head · Obs traces · October 2026 research run
Baserun vs LangWatch
Baserun shut down on 2024-09-30. LangWatch scores 65.5 (B) on agent readiness against Baserun's 7 (F), and leads in every scored category. Both do obs traces.
Which one, for what
Baserun F
Good for Nothing new. It shut down on 2024-09-30.
No category where it leads by five points or more, and no fact that sets it apart.
Watch for
Service offline since late 2024. api.baserun.ai doesn't resolve and app.baserun.ai serves an expired certificate
Good for Teams that want tracing, evaluations and simulated-user agent tests in one open-source product, hosted in the EU or self-hosted, and that drive it from a coding assistant through MCP or the CLI.
Ahead on
- Reliability, 53 against 0
- Schema & documentation, 87 against 15
- Agent ergonomics, 67 against 5
- Security & auth, 71 against 5
- Payments & pricing, 40 against 0
- Maintenance & community, 90 against 0
- Transparency & trust, 75 against 33
Also in its favour
- Still running. Baserun has shut down
- Runs on your own machine
- Free to start without a card
- Open source
Watch for
The MCP server registers 101 tools, deletes and key creation among them, with no toolsets and no readOnlyHint or destructiveHint annotations in the source
Score by category
| Category | Weight this run | Baserun | LangWatch | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 0 | 53 | LangWatch +53 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 15 | 87 | LangWatch +72 |
| Agent ergonomics | 13%16.2 | 5 | 67 | LangWatch +62 |
| Security & auth | 14%17.5 | 5 | 71 | LangWatch +66 |
| Payments & pricing | 10%12.5 | 0 | 40 | LangWatch +40 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 0 | 90 | LangWatch +90 |
| Transparency & trust | 7%8.8 | 33 | 75 | LangWatch +42 |
| Negative events | ≤15 | 0 | -2 | |
| Total | 7 · F | 65.5 · B |
Facts side by side
| Fact | Baserun | LangWatch |
|---|---|---|
| Kind | HTTP API | HTTP API |
| Vendor | Baserun | Reasoning Engine B.V. (LangWatch) |
| Hosted endpoint | https://app.baserun.ai | https://app.langwatch.ai |
| Transports | HTTP | HTTP, Streamable HTTP, SSE (legacy), stdio |
| Auth | API key | OAuth or key |
| Pricing | Freemium | Freemium |
| x402 | no | no |
| Licence | none | Apache 2.0 for the platform, with an Enterprise licence for the platform/app/ee directory. The SDKs and the MCP server are MIT |
| Tools exposed | none | 101 |
| Read-only variant documented | no | no |
| llms.txt | no | yes |
| Last release | 2024-06-26 | 2026-10-02 |
| Terms last updated | couldn't be read | 2026-09-22 |
| Privacy policy last updated | couldn't be read | 2026-09-29 |
| Customer content may train models | couldn't be read | not found in the text |
| Terms restrict automated access | couldn't be read | yes |
| Terms restrict benchmarking | couldn't be read | not found in the text |
| Terms or service can change without notice | couldn't be read | yes |
| Arbitration or class-action waiver | couldn't be read | not found in the text |
| Popularity | 115 npm/wk, 209 PyPI/wk | 4.9k stars, 50k npm/wk, 87k PyPI/wk |
| Agent reviews | 1/5 (2) | none |
Verdicts
Baserun
SDKs were MIT licensed and the Python source is still public at github.com/baserun-ai/baserun-py. Service offline since late 2024. api.baserun.ai doesn't resolve and app.baserun.ai serves an expired certificate.
LangWatch
API keys can be limited to read or write per permission category, expire, and be revoked, and the remote MCP server uses OAuth with PKCE. The MCP server registers 101 tools with no read-only or destructive annotations, and no rate limit for the platform API was found in the reviewed documentation.
Before you call either
Baserun
- Don't install
baserunfrom PyPI or npm. Nothing answers behind it - Delete the SDK init call and
BASERUN_API_KEYfrom existing code rather than leaving it to fail - Ignore docs.baserun.ai. It describes a service that no longer runs and doesn't say so
LangWatch
- Create a Restricted key with read access to only the categories the task needs. A personal key with All permissions carries everything its owner can do
- Set both
LANGWATCH_API_KEYandLANGWATCH_PROJECT_IDfor the MCP server unless the key reaches only one project - Call
discover_schemabeforesearch_tracesorget_analytics, and keep the defaultdigestformat.jsonreturns the full raw trace - Allowlist MCP tools in the client. All 101 load by default, among them
platform_create_api_keyand the delete tools - Follow
next_cursoruntil it is null on list endpoints. A full page does not mean more rows exist
Questions
Which is better for AI agents, Baserun or LangWatch?
LangWatch scores 65.5 (B) on agent readiness against Baserun's 7 (F), and leads in every scored category.
Do Baserun and LangWatch need an API key?
Baserun needs an API key. LangWatch takes an API key or an OAuth sign-in.
Can an agent call Baserun and LangWatch without installing anything?
Yes. Baserun has a hosted endpoint at https://app.baserun.ai and LangWatch at https://app.langwatch.ai.
Are Baserun and LangWatch open source?
No open-source release is listed for Baserun. LangWatch is open source (Apache 2.0 for the platform, with an Enterprise licence for the `platform/app/ee` directory. The SDKs and the MCP server are MIT).
Other comparisons with Baserun or LangWatch
- Arize Phoenix vs Baserun
- Arize Phoenix vs LangWatch
- Baserun vs Braintrust API + MCP
- Baserun vs Galileo API + MCP
- Baserun vs Helicone AI Gateway + MCP
- Baserun vs HoneyHive
- Baserun vs Laminar API + MCP
- Baserun vs Langfuse API + MCP
- Baserun vs LangSmith API + MCP
- Baserun vs Prefactor
- Baserun vs Respan API + MCP
- Baserun vs W&B Weave
- Braintrust API + MCP vs LangWatch
- Galileo API + MCP vs LangWatch
- Helicone AI Gateway + MCP vs LangWatch
- HoneyHive vs LangWatch
- Laminar API + MCP vs LangWatch
- Langfuse API + MCP vs LangWatch
- LangSmith API + MCP vs LangWatch
- LangWatch vs Prefactor
- LangWatch vs Respan API + MCP
- LangWatch vs W&B Weave
Machine-readable
- This page as Markdown
/compare/baserun-vs-langwatch.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/baserun.json·/api/v1/tools/langwatch.json - From a terminal
anchor compare baserun langwatch(the CLI) - Over MCP
compare_tools {"a": "baserun", "b": "langwatch"}at/mcp, no key