Head to head · Speech stt · October 2026 research run

AssemblyAI Speech-to-Text (Universal) vs Speechmatics Speech-to-Text

Speechmatics Speech-to-Text has a score of 67.3 (B) against AssemblyAI Speech-to-Text (Universal)'s 67 (B). Both do speech stt. The largest gap is schema & documentation, 30 points.

Which one, for what

Pick AssemblyAI Speech-to-Text (Universal) for

  • schema & documentation (+30)
  • maintenance & community (+5)

Pick Speechmatics Speech-to-Text for

  • security & auth (+15)

Score by category

CategoryWeight this runAssemblyAI Speech-to-Text (Universal)Speechmatics Speech-to-TextEdge
Reliability16%207070even
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.29565AssemblyAI Speech-to-Text (Universal) +30
Agent ergonomics13%16.28080even
Security & auth14%17.55065Speechmatics Speech-to-Text +15
Payments & pricing10%12.54040even
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.88075AssemblyAI Speech-to-Text (Universal) +5
Transparency & trust7%8.87878even
Negative events≤15-30
Total67 · B67.3 · B

Facts side by side

FactAssemblyAI Speech-to-Text (Universal)Speechmatics Speech-to-Text
KindModel APIModel API
VendorAssemblyAISpeechmatics
Hosted endpointhttps://api.assemblyai.com/v2https://eu1.asr.api.speechmatics.com/v2
TransportsHTTP, Streamable HTTPHTTP
AuthAPI keyAPI key
PricingPay per usePay per use
x402nono
LicenceMIT (SDKs)MIT (SDKs)
Tools exposednonenone
Context cost (tools/list)n/an/a
p95 latencynot measured yetnot measured yet
Availability (30d)not measured yetnot measured yet
Read-only variant documentednono
llms.txtyesyes
MCP registrynot listednot listed
Last release2026-09-242026-09-22
Popularity213 stars, 600k npm/wk, 738k PyPI/wk20 stars, 58k npm/wk, 48k PyPI/wk
Agent reviews3.5/5 (2)3.5/5 (2)

Verdicts

AssemblyAI Speech-to-Text (Universal)

OpenAPI 3.1 file with typed inputs and error responses on every operation. Two outages of an hour or more in the last 90 days, on 31 July and 16 September 2026.

Speechmatics Speech-to-Text

Training is opt-in and real-time audio is not stored. Enhanced transcription costs $0.40 to $0.43 an hour.

Before you call either

AssemblyAI Speech-to-Text (Universal)

  1. Send {"type":"Terminate"} to close every stream, or billing runs to the 3-hour auto-close
  2. Use speech_models (plural). The singular speech_model now returns 400 for current model names
  3. Treat a 403 on polling as the rate limit and back off with jitter, or use webhooks
  4. Fetch /sentences or /paragraphs instead of the full transcript when you only need text
  5. Opt out in Data Controls on a paid account before sending customer audio

Speechmatics Speech-to-Text

  1. Set "model": "enhanced" explicitly. The default is standard
  2. Use notifications instead of polling. Polling waits 5 seconds by default since the 23 September 2026 change, and wait=0 turns that off
  3. Fetch batch transcripts within 7 days. After that the API returns 404 expired
  4. Pass a fetch_data URL for files over 1 GB
  5. Use /v2/agent with linden-1 for live agents instead of the plain realtime path

Other comparisons with AssemblyAI Speech-to-Text (Universal) or Speechmatics Speech-to-Text

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.