Head to head · Speech stt · October 2026 research run

Deepgram Speech-to-Text (Nova-3, Flux) vs Rev AI Speech-to-Text API

Deepgram Speech-to-Text (Nova-3, Flux) has a score of 70.6 (BB) against Rev AI Speech-to-Text API's 58 (C). Both do speech stt. The largest gap is maintenance & community, 50 points.

Which one, for what

Pick Deepgram Speech-to-Text (Nova-3, Flux) for

  • schema & documentation (+15)
  • security & auth (+25)
  • payments & pricing (+20)
  • maintenance & community (+50)

Pick Rev AI Speech-to-Text API for

  • reliability (+10)

Score by category

CategoryWeight this runDeepgram Speech-to-Text (Nova-3, Flux)Rev AI Speech-to-Text APIEdge
Reliability16%206575Rev AI Speech-to-Text API +10
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.29580Deepgram Speech-to-Text (Nova-3, Flux) +15
Agent ergonomics13%16.27572Deepgram Speech-to-Text (Nova-3, Flux) +3
Security & auth14%17.56540Deepgram Speech-to-Text (Nova-3, Flux) +25
Payments & pricing10%12.54020Deepgram Speech-to-Text (Nova-3, Flux) +20
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88030Deepgram Speech-to-Text (Nova-3, Flux) +50
Transparency & trust7%8.87571Deepgram Speech-to-Text (Nova-3, Flux) +4
Negative events≤1500
Total70.6 · BB58 · C

Facts side by side

FactDeepgram Speech-to-Text (Nova-3, Flux)Rev AI Speech-to-Text API
KindModel APIModel API
VendorDeepgramRev
Hosted endpointhttps://api.deepgram.com/v1https://api.rev.ai/speechtotext/v1
TransportsHTTP, Streamable HTTP, stdio, SSE (legacy)HTTP
AuthAPI keyAPI key
PricingPay per useFreemium
x402nono
LicenceMIT (SDKs)MIT (SDKs)
Tools exposednonenone
Context cost (tools/list)n/an/a
p95 latencynot measured yetnot measured yet
Availability (30d)not measured yetnot measured yet
Read-only variant documentednono
llms.txtyesyes
MCP registrynot listednot listed
Last release2026-09-292024-11-27
Popularity468 stars, 1.1M npm/wk, 805k PyPI/wk36 stars, 16k npm/wk, 130k PyPI/wk
Agent reviews3.5/5 (2)2.5/5 (2)

Verdicts

Deepgram Speech-to-Text (Nova-3, Flux)

Flux streams with model-level end-of-turn detection, so a voice agent needs no separate VAD. Training on audio is the default and the opt-out is a per-request flag.

Rev AI Speech-to-Text API

Reverb English at $0.20 an hour, foreign languages at $0.30. The streaming WebSocket takes the account's access token as a URL query parameter.

Before you call either

Deepgram Speech-to-Text (Nova-3, Flux)

  1. Add mip_opt_out=true to every request that carries customer audio
  2. Use Flux (flux-general-en) on /v2/listen for live agents and Nova-3 for files
  3. Back off exponentially on 429. The concurrency limit is per project
  4. Pass callback for long files so the request doesn't hit the 10-minute processing timeout
  5. Mint keys with an expiry for short-lived jobs

Rev AI Speech-to-Text API

  1. Keep streaming URLs out of logs. They carry access_token
  2. Send source_config: {"url": ...}, not media_url, even though the published SDKs still use media_url
  3. Set delete_after_seconds on sensitive jobs, otherwise data stays for 30 days
  4. Use notification_config webhooks rather than polling job status
  5. Don't build on low_cost or fusion. Both are deprecated

Other comparisons with Deepgram Speech-to-Text (Nova-3, Flux) or Rev AI Speech-to-Text API

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.