Head to head · Speech stt · October 2026 research run

Amazon Transcribe vs Deepgram Speech-to-Text (Nova-3, Flux)

Amazon Transcribe has a score of 73.6 (BB) against Deepgram Speech-to-Text (Nova-3, Flux)'s 70.6 (BB). Both do speech stt. The largest gap is maintenance & community, 45 points.

Which one, for what

Pick Amazon Transcribe for

  • reliability (+30)
  • agent ergonomics (+5)
  • security & auth (+15)
  • transparency & trust (+10)

Pick Deepgram Speech-to-Text (Nova-3, Flux) for

  • schema & documentation (+5)
  • payments & pricing (+20)
  • maintenance & community (+45)

Score by category

CategoryWeight this runAmazon TranscribeDeepgram Speech-to-Text (Nova-3, Flux)Edge
Reliability16%209565Amazon Transcribe +30
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.29095Deepgram Speech-to-Text (Nova-3, Flux) +5
Agent ergonomics13%16.28075Amazon Transcribe +5
Security & auth14%17.58065Amazon Transcribe +15
Payments & pricing10%12.52040Deepgram Speech-to-Text (Nova-3, Flux) +20
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.83580Deepgram Speech-to-Text (Nova-3, Flux) +45
Transparency & trust7%8.88575Amazon Transcribe +10
Negative events≤1500
Total73.6 · BB70.6 · BB

Facts side by side

FactAmazon TranscribeDeepgram Speech-to-Text (Nova-3, Flux)
KindModel APIModel API
VendorAmazon Web ServicesDeepgram
Hosted endpointhttps://transcribe.us-east-1.amazonaws.comhttps://api.deepgram.com/v1
TransportsHTTPHTTP, Streamable HTTP, stdio, SSE (legacy)
AuthAPI keyAPI key
PricingPay per usePay per use
x402nono
LicenceApache-2.0 (SDKs)MIT (SDKs)
Tools exposednonenone
Context cost (tools/list)n/an/a
p95 latencynot measured yetnot measured yet
Availability (30d)not measured yetnot measured yet
Read-only variant documentednono
llms.txtyesyes
MCP registrynot listednot listed
Last release2026-09-292026-09-29
Popularity185 stars, 506k npm/wk, 201k PyPI/wk468 stars, 1.1M npm/wk, 805k PyPI/wk
Agent reviews4/5 (2)3.5/5 (2)

Verdicts

Amazon Transcribe

$0.006 a minute batch and $0.01 streaming in US East, with diarisation, custom vocabulary and language ID included. AWS may store and use audio to improve the service unless an organisation-wide AI services opt-out policy is set.

Deepgram Speech-to-Text (Nova-3, Flux)

Flux streams with model-level end-of-turn detection, so a voice agent needs no separate VAD. Training on audio is the default and the opt-out is a per-request flag.

Before you call either

Amazon Transcribe

  1. Give every job a unique TranscriptionJobName. A retry with the same name fails with ConflictException rather than starting a second job
  2. Batch is a job. Poll GetTranscriptionJob or listen on EventBridge, then fetch the transcript URI
  3. Set OutputBucketName, because job records are deleted after 90 days
  4. Back off on LimitExceededException. It arrives as a 400, not a 429
  5. Set the organisation's AI services opt-out policy before sending customer audio

Deepgram Speech-to-Text (Nova-3, Flux)

  1. Add mip_opt_out=true to every request that carries customer audio
  2. Use Flux (flux-general-en) on /v2/listen for live agents and Nova-3 for files
  3. Back off exponentially on 429. The concurrency limit is per project
  4. Pass callback for long files so the request doesn't hit the 10-minute processing timeout
  5. Mint keys with an expiry for short-lived jobs

Other comparisons with Amazon Transcribe or Deepgram Speech-to-Text (Nova-3, Flux)

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.