Category · Work & business apps

Product analytics and experimentation tools for AI agents

Tools that record how people use a product and answer questions about funnels, retention and cohorts, with feature flags and experiments where the vendor has them. Compared on query access, event capture, export and experiment results.

Capability keys analytics.query · analytics.events · analytics.funnels · analytics.experiments · analytics.flags · All tools

letme.dev/analytics.query picks the top-graded tool in this list and says how to call it direct; calling through letme comes later.

5listings graded
0agent-ready (BB+)
0desk reviews by the panel
0accept x402
8 Oct 16:50last updated (UTC)
Filters
Grade
Agent rating
Where it runs
Auth
Pricing
Status
5 tools
Compare#ToolCategoryGradeScoreAgent ratingPrice / x402Details
151 PostHogPostHog Inc. · HTTP API Product analytics B 68.4 none $0.0001 / tx
199 AmplitudeAmplitude, Inc. · HTTP API Product analytics B 66.2 none Freemium
281 MixpanelMixpanel, Inc. · HTTP API Product analytics B 62 none $140 / mo
442 HeapContentsquare (Content Square, Inc.) · HTTP API Product analytics D 53 none Freemium
491 PendoPendo.io, Inc. · HTTP API Product analytics D 48.2 none Paid

p95 latency and context cost come from our probes, which haven't run yet, so those columns start hidden. Grades run from AA to F, and agent-ready means BB or better. Filters, sorting and export run in your browser; the table is complete without JavaScript.

How we test this category

One fixed event set loaded into each listing, then the same questions asked through its API or MCP server (a funnel conversion, a retention curve, a cohort size and an experiment result). We check the answers against known values and the limits on queries. In this run listings are graded from public evidence against the published checklist. This test hasn't run yet, so Task success is pending and the grades here come from the categories assessed from public evidence.

How the ranking works

Every listing is scored 0 to 100 and given a grade from AA to F. In the October 2026 research run, 7 of the 9 weighted categories are scored from public evidence (status history, docs, pricing, terms, source and security pages) against a published checklist, with the reason and sources for every score on the listing. Performance and Task success wait for our probes and task suites, so their weight is shared across the rest until they run. Negative events deduct up to 15 points. Read the methodology.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.