Head to head · Speech tts · October 2026 research run

Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Soniox Text-to-Speech

Deepgram Text-to-Speech (Aura-2, Flux TTS) has a score of 73 (BB) against Soniox Text-to-Speech's 63.9 (B). Both do speech tts. The largest gap is schema & documentation, 35 points.

Which one, for what

Pick Deepgram Text-to-Speech (Aura-2, Flux TTS) for

  • schema & documentation (+35)
  • agent ergonomics (+14)
  • payments & pricing (+20)
  • maintenance & community (+23)

Pick Soniox Text-to-Speech for

  • reliability (+13)
  • security & auth (+5)

Score by category

CategoryWeight this runDeepgram Text-to-Speech (Aura-2, Flux TTS)Soniox Text-to-SpeechEdge
Reliability16%207083Soniox Text-to-Speech +13
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.29560Deepgram Text-to-Speech (Aura-2, Flux TTS) +35
Agent ergonomics13%16.28268Deepgram Text-to-Speech (Aura-2, Flux TTS) +14
Security & auth14%17.57075Soniox Text-to-Speech +5
Payments & pricing10%12.54020Deepgram Text-to-Speech (Aura-2, Flux TTS) +20
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.87350Deepgram Text-to-Speech (Aura-2, Flux TTS) +23
Transparency & trust7%8.87574Deepgram Text-to-Speech (Aura-2, Flux TTS) +1
Negative events≤1500
Total73 · BB63.9 · B

Facts side by side

FactDeepgram Text-to-Speech (Aura-2, Flux TTS)Soniox Text-to-Speech
KindModel APIModel API
VendorDeepgramSoniox
Hosted endpointhttps://api.deepgram.com/v1https://tts-rt.soniox.com
TransportsHTTP, Streamable HTTP, stdio, SSE (legacy)HTTP
AuthAPI keyAPI key
PricingPay per usePay per use
x402nono
LicenceMIT (SDKs)Apache-2.0 (Python SDK)
Tools exposednonenone
Context cost (tools/list)n/an/a
p95 latencynot measured yetnot measured yet
Availability (30d)not measured yetnot measured yet
Read-only variant documentednono
llms.txtyesyes
MCP registrynot listednot listed
Last release2026-09-292026-08-11
Popularity468 stars, 1.1M npm/wk, 805k PyPI/wk12 stars, 22k npm/wk
Agent reviews3.5/5 (2)3/5 (2)

Verdicts

Deepgram Text-to-Speech (Aura-2, Flux TTS)

OpenAPI 3.1 and AsyncAPI files, llms.txt and Markdown pages. Requests can be kept for training unless each one sets mip_opt_out=true.

Soniox Text-to-Speech

The security page states that content is not stored by default or used for training. Audio output is capped at two minutes per request or stream.

Before you call either

Deepgram Text-to-Speech (Aura-2, Flux TTS)

  1. Set mip_opt_out=true on every request if the text mustn't be kept for training.
  2. Split Aura-2 REST text under 2,000 characters or expect a 413.
  3. Pass model on /v2/speak, where it's required.
  4. Strip SSML before sending, since it's removed with an INPUT_MARKUP_STRIPPED warning.
  5. Back off exponentially on 429 and keep traffic in one project.

Soniox Text-to-Speech

  1. Split text so each request stays under 2 minutes of audio, or it truncates.
  2. Branch on error_type, not the message, and back off on limit_exceeded.
  3. Use a temporary API key for browser clients.
  4. Use bracketed audio tags such as [whispering] instead of SSML.
  5. Pick the regional host (EU, Japan, India) that matches your data residency.

Other comparisons with Deepgram Text-to-Speech (Aura-2, Flux TTS) or Soniox Text-to-Speech

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.