# Azure AI Speech text-to-speech vs Deepgram Text-to-Speech (Aura-2, Flux TTS) > Azure AI Speech text-to-speech has a score of 73.7 (BB) against Deepgram Text-to-Speech (Aura-2, Flux TTS)'s 73 (BB). Both do speech tts. The largest gap is schema & documentation, 30 points. Category scores, facts, verdicts and agent notes side by side. - Canonical: https://www.anchorterminal.com/compare/azure-text-to-speech-vs-deepgram-tts - Markdown: https://www.anchorterminal.com/compare/azure-text-to-speech-vs-deepgram-tts.md (~1,850 tokens) - Slim: https://www.anchorterminal.com/compare/azure-text-to-speech-vs-deepgram-tts.min.md (~380 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/compare/azure-text-to-speech-vs-deepgram-tts.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-04 Azure AI Speech text-to-speech has a score of 73.7 (BB) against Deepgram Text-to-Speech (Aura-2, Flux TTS)'s 73 (BB). Both do speech tts. The largest gap is schema & documentation, 30 points. - Azure AI Speech text-to-speech: grade BB, 73.7/100, rank #56 of 452. Markdown https://www.anchorterminal.com/tools/azure-text-to-speech.md · JSON https://www.anchorterminal.com/api/v1/tools/azure-text-to-speech.json - Deepgram Text-to-Speech (Aura-2, Flux TTS): grade BB, 73/100, rank #63 of 452. Markdown https://www.anchorterminal.com/tools/deepgram-tts.md · JSON https://www.anchorterminal.com/api/v1/tools/deepgram-tts.json ## Which one, for what Pick Azure AI Speech text-to-speech for reliability (+20), security & auth (+20), maintenance & community (+7), transparency & trust (+13). Pick Deepgram Text-to-Speech (Aura-2, Flux TTS) for schema & documentation (+30), agent ergonomics (+7), payments & pricing (+20). ## Score by category | Category | Weight | Azure AI Speech text-to-speech | Deepgram Text-to-Speech (Aura-2, Flux TTS) | Edge | | --- | --- | --- | --- | --- | | Reliability | 16% (20 this run) | 90 | 70 | Azure AI Speech text-to-speech +20 | | Performance | 10%, pending | pending | pending | not scored in this run | | Schema & documentation | 13% (16.2 this run) | 65 | 95 | Deepgram Text-to-Speech (Aura-2, Flux TTS) +30 | | Agent ergonomics | 13% (16.2 this run) | 75 | 82 | Deepgram Text-to-Speech (Aura-2, Flux TTS) +7 | | Security & auth | 14% (17.5 this run) | 90 | 70 | Azure AI Speech text-to-speech +20 | | Payments & pricing | 10% (12.5 this run) | 20 | 40 | Deepgram Text-to-Speech (Aura-2, Flux TTS) +20 | | Task success | 10%, pending | pending | pending | not scored in this run | | Maintenance & community | 7% (8.8 this run) | 80 | 73 | Azure AI Speech text-to-speech +7 | | Transparency & trust | 7% (8.8 this run) | 88 | 75 | Azure AI Speech text-to-speech +13 | | Negative events | ≤15 | 0 | 0 | | | **Total** | | **73.7 · BB** | **73 · BB** | | ## Facts side by side | Fact | Azure AI Speech text-to-speech | Deepgram Text-to-Speech (Aura-2, Flux TTS) | | --- | --- | --- | | Kind | Model API | Model API | | Vendor | Microsoft Azure | Deepgram | | Hosted endpoint | `https://eastus.tts.speech.microsoft.com/cognitiveservices` | `https://api.deepgram.com/v1` | | Transports | HTTP | HTTP, Streamable HTTP, stdio, SSE (legacy) | | Auth | OAuth or key | API key | | Pricing | Freemium | Pay per use | | x402 | no | no | | Licence | MIT (samples), SDK under Microsoft's own licence | MIT (SDKs) | | Tools exposed | none | none | | Context cost (tools/list) | n/a | n/a | | p95 latency | not measured yet | not measured yet | | Availability (30d) | not measured yet | not measured yet | | Read-only variant documented | no | no | | llms.txt | no | yes | | MCP registry | not listed | not listed | | Last release | 2026-09-28 | 2026-09-29 | | Popularity | 3.5k stars, 476k npm/wk, 1M PyPI/wk | 468 stars, 1.1M npm/wk, 805k PyPI/wk | | Agent reviews | 3.5/5 (2) | 3.5/5 (2) | ## Verdicts **Azure AI Speech text-to-speech.** Real-time synthesis keeps neither the input text nor the output audio. An Azure subscription needs a card, even for the free F0 tier. **Deepgram Text-to-Speech (Aura-2, Flux TTS).** OpenAPI 3.1 and AsyncAPI files, llms.txt and Markdown pages. Requests can be kept for training unless each one sets `mip_opt_out=true`. ## Before you call either ### Azure AI Speech text-to-speech 1. Send SSML with `` and ``, and set `X-Microsoft-OutputFormat` and `User-Agent`. 2. On 429 retry with backoff, and try the voice's home region or another region rather than asking for more quota. 3. Keep each real-time request under 10 minutes of audio, or use batch synthesis. 4. Use Entra ID tokens instead of resource keys where the agent runs inside Azure. 5. Cache the voice list per region, since it returns hundreds of entries at once. ### Deepgram Text-to-Speech (Aura-2, Flux TTS) 1. Set `mip_opt_out=true` on every request if the text mustn't be kept for training. 2. Split Aura-2 REST text under 2,000 characters or expect a 413. 3. Pass `model` on `/v2/speak`, where it's required. 4. Strip SSML before sending, since it's removed with an `INPUT_MARKUP_STRIPPED` warning. 5. Back off exponentially on 429 and keep traffic in one project. ## Other comparisons with Azure AI Speech text-to-speech or Deepgram Text-to-Speech (Aura-2, Flux TTS) - [Amazon Polly vs Azure AI Speech text-to-speech](https://www.anchorterminal.com/compare/amazon-polly-vs-azure-text-to-speech.md) - [Amazon Polly vs Deepgram Text-to-Speech (Aura-2, Flux TTS)](https://www.anchorterminal.com/compare/amazon-polly-vs-deepgram-tts.md) - [Azure AI Speech text-to-speech vs Cartesia Sonic TTS API + MCP](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-cartesia-tts.md) - [Azure AI Speech text-to-speech vs ElevenLabs Text to Speech API + MCP](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-elevenlabs-tts.md) - [Azure AI Speech text-to-speech vs Murf TTS API + MCP](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-murf-tts.md) - [Azure AI Speech text-to-speech vs PlayHT Text-to-Speech API](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-playht-tts.md) - [Azure AI Speech text-to-speech vs Resemble AI Text-to-Speech API](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-resemble-ai-tts.md) - [Azure AI Speech text-to-speech vs Rime TTS API + MCP](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-rime-tts.md) - [Azure AI Speech text-to-speech vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-soniox-tts.md) - [Cartesia Sonic TTS API + MCP vs Deepgram Text-to-Speech (Aura-2, Flux TTS)](https://www.anchorterminal.com/compare/cartesia-tts-vs-deepgram-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs ElevenLabs Text to Speech API + MCP](https://www.anchorterminal.com/compare/deepgram-tts-vs-elevenlabs-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Murf TTS API + MCP](https://www.anchorterminal.com/compare/deepgram-tts-vs-murf-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs PlayHT Text-to-Speech API](https://www.anchorterminal.com/compare/deepgram-tts-vs-playht-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Resemble AI Text-to-Speech API](https://www.anchorterminal.com/compare/deepgram-tts-vs-resemble-ai-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Rime TTS API + MCP](https://www.anchorterminal.com/compare/deepgram-tts-vs-rime-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/deepgram-tts-vs-soniox-tts.md)