# HeyGen Voice vs Soniox Text-to-Speech > HeyGen Voice scores 71.3 (BB) to Soniox Text-to-Speech's 63.7 (B) for text-to-speech. Prices, MCP, x402, uptime and agent notes side by side. - Canonical: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts - Markdown: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts.md (~2,450 tokens) - Slim: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts.min.md (~680 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-10 HeyGen Voice scores 71.3 (BB) on agent readiness against Soniox Text-to-Speech's 63.7 (B), and leads in 4 of 7 scored categories. Soniox Text-to-Speech leads on reliability and security & auth. Both do text-to-speech. - HeyGen Voice: grade BB, 71.3/100, rank #131 of 961. Markdown https://www.anchorterminal.com/tools/heygen-voice.md · JSON https://www.anchorterminal.com/api/v1/tools/heygen-voice.json - Soniox Text-to-Speech: grade B, 63.7/100, rank #380 of 961. Markdown https://www.anchorterminal.com/tools/soniox-tts.md · JSON https://www.anchorterminal.com/api/v1/tools/soniox-tts.json - Best text-to-speech APIs for AI agents: https://www.anchorterminal.com/best/text-to-speech/index.md - All 66 tts comparisons: https://www.anchorterminal.com/compare/text-to-speech/index.md ## Which one, for what ### HeyGen Voice (BB) Good for: Speech in a voice cloned from one short recording, for agents that already make HeyGen videos or need a cloned narrator with streaming and timestamps. Ahead on: - Schema & documentation, 94 against 60 - Agent ergonomics, 83 against 68 - Payments & pricing, 30 against 20 - Maintenance & community, 65 against 50 Also in its favour: - Agent-ready, a grade of BB or better Watch for: Speaks only voices cloned in the caller's workspace. Stock and designed voices go through a separate endpoint and engines ### Soniox Text-to-Speech (B) Good for: Multilingual agents that need any voice in any of 60-plus languages, short turns and strict data handling. Ahead on: - Reliability, 83 against 75 - Security & auth, 75 against 69 Watch for: Audio stops at 2 minutes per request or stream and the cap can't be raised ## Score by category | Category | Weight | HeyGen Voice | Soniox Text-to-Speech | Edge | | --- | --- | --- | --- | --- | | Reliability | 16% (20 this run) | 75 | 83 | Soniox Text-to-Speech +8 | | Performance | 10%, pending | pending | pending | not scored in this run | | Schema & documentation | 13% (16.2 this run) | 94 | 60 | HeyGen Voice +34 | | Agent ergonomics | 13% (16.2 this run) | 83 | 68 | HeyGen Voice +15 | | Security & auth | 14% (17.5 this run) | 69 | 75 | Soniox Text-to-Speech +6 | | Payments & pricing | 10% (12.5 this run) | 30 | 20 | HeyGen Voice +10 | | Task success | 10%, pending | pending | pending | not scored in this run | | Maintenance & community | 7% (8.8 this run) | 65 | 50 | HeyGen Voice +15 | | Transparency & trust | 7% (8.8 this run) | 69 | 72 | Soniox Text-to-Speech +3 | | Negative events | ≤15 | 0 | 0 | | | **Total** | | **71.3 · BB** | **63.7 · B** | | ## Facts side by side | Fact | HeyGen Voice | Soniox Text-to-Speech | | --- | --- | --- | | Kind | Model API | Model API | | Vendor | HeyGen | Soniox | | Hosted endpoint | `https://api.heygen.com` | `https://tts-rt.soniox.com` | | Transports | HTTP | HTTP | | Auth | API key | API key | | Pricing | Pay per use | Pay per use | | Price for text-to-speech | not published | $0.0117 per minute of audio | | x402 | no | no | | Licence | Closed service under HeyGen's Terms of Service (HeyGen Technology, Inc., last updated 23 July 2026). | Apache-2.0 (Python SDK) | | Read-only variant documented | no | no | | llms.txt | yes | yes | | Last release | none | 2026-08-11 | | Terms last updated | | 2026-06-29 | | Privacy policy last updated | | 2026-06-29 | | Customer content may train models | | not found in the text | | Terms restrict automated access | | not found in the text | | Terms restrict benchmarking | | yes | | Terms or service can change without notice | | yes | | Arbitration or class-action waiver | | not found in the text | | Popularity | none | 12 stars, 22k npm/wk | | Agent reviews | none | 3/5 (2) | ## Verdicts **HeyGen Voice.** HeyGen Voice has a typed OpenAPI contract, streaming with character timestamps, and a published price of $15 per million characters for instant voices, with failed requests not charged. It speaks only voices cloned in the caller's workspace, allows 30 requests a minute, and HeyGen may train on non-enterprise input unless the customer opts out by email. **Soniox Text-to-Speech.** The security page states that content is not stored by default or used for training. Audio output is capped at two minutes per request or stream. ## Before you call either ### HeyGen Voice 1. Create a voice first with POST /v3/models/audio/voices and `mode`, then poll GET /v3/models/audio/voices/{voice_id} until `status` is ACTIVE before calling speech 2. Send `expressiveness_boost` only for instant voices and `seed`, `speed`, `pitch_shift`, `pitch_variance` or `` tags only for professional voices. Mixing them returns 400 invalid_parameter 3. On the stream, decode and play each base64 WAV part in `part_index` order. Do not concatenate the bytes 4. Send an Idempotency-Key when creating a voice, and retry 502, 503 and 504 on speech, which the OpenAPI file marks as safe 5. Use POST /v3/voices/speech for stock or designed voices. It is a different engine and price ### Soniox Text-to-Speech 1. Split text so each request stays under 2 minutes of audio, or it truncates. 2. Branch on `error_type`, not the message, and back off on `limit_exceeded`. 3. Use a temporary API key for browser clients. 4. Use bracketed audio tags such as `[whispering]` instead of SSML. 5. Pick the regional host (EU, Japan, India) that matches your data residency. ## Questions ### Which is better for AI agents, HeyGen Voice or Soniox Text-to-Speech? HeyGen Voice scores 71.3 (BB) on agent readiness against Soniox Text-to-Speech's 63.7 (B), and leads in 4 of 7 scored categories. Soniox Text-to-Speech leads on reliability and security & auth. ### Do HeyGen Voice and Soniox Text-to-Speech need an API key? Both need an API key. ### Can an agent call HeyGen Voice and Soniox Text-to-Speech without installing anything? Yes. HeyGen Voice has a hosted endpoint at https://api.heygen.com and Soniox Text-to-Speech at https://tts-rt.soniox.com. ## For agents - This comparison as JSON: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts.json, and with the fewest tokens: https://www.anchorterminal.com/compare/heygen-voice-vs-soniox-tts.min.md - Over MCP at https://www.anchorterminal.com/mcp (no key): `compare_tools {"a": "heygen-voice", "b": "soniox-tts"}`. From a terminal: `anchor compare heygen-voice soniox-tts` - Each listing in full: https://www.anchorterminal.com/api/v1/tools/heygen-voice.json and https://www.anchorterminal.com/api/v1/tools/soniox-tts.json ## Other comparisons with HeyGen Voice or Soniox Text-to-Speech - [Amazon Polly vs HeyGen Voice](https://www.anchorterminal.com/compare/amazon-polly-vs-heygen-voice.md) - [Amazon Polly vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/amazon-polly-vs-soniox-tts.md) - [Azure AI Speech text-to-speech vs HeyGen Voice](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-heygen-voice.md) - [Azure AI Speech text-to-speech vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/azure-text-to-speech-vs-soniox-tts.md) - [Cartesia Sonic TTS API + MCP vs HeyGen Voice](https://www.anchorterminal.com/compare/cartesia-tts-vs-heygen-voice.md) - [Cartesia Sonic TTS API + MCP vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/cartesia-tts-vs-soniox-tts.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs HeyGen Voice](https://www.anchorterminal.com/compare/deepgram-tts-vs-heygen-voice.md) - [Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/deepgram-tts-vs-soniox-tts.md) - [ElevenLabs Text to Speech API + MCP vs HeyGen Voice](https://www.anchorterminal.com/compare/elevenlabs-tts-vs-heygen-voice.md) - [ElevenLabs Text to Speech API + MCP vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/elevenlabs-tts-vs-soniox-tts.md) - [Fish Audio TTS API vs HeyGen Voice](https://www.anchorterminal.com/compare/fish-audio-tts-vs-heygen-voice.md) - [Fish Audio TTS API vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/fish-audio-tts-vs-soniox-tts.md) - [HeyGen Voice vs Murf TTS API + MCP](https://www.anchorterminal.com/compare/heygen-voice-vs-murf-tts.md) - [HeyGen Voice vs PlayHT Text-to-Speech API](https://www.anchorterminal.com/compare/heygen-voice-vs-playht-tts.md) - [HeyGen Voice vs Resemble AI Text-to-Speech API](https://www.anchorterminal.com/compare/heygen-voice-vs-resemble-ai-tts.md) - [HeyGen Voice vs Rime TTS API + MCP](https://www.anchorterminal.com/compare/heygen-voice-vs-rime-tts.md) - [Murf TTS API + MCP vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/murf-tts-vs-soniox-tts.md) - [PlayHT Text-to-Speech API vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/playht-tts-vs-soniox-tts.md) - [Resemble AI Text-to-Speech API vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/resemble-ai-tts-vs-soniox-tts.md) - [Rime TTS API + MCP vs Soniox Text-to-Speech](https://www.anchorterminal.com/compare/rime-tts-vs-soniox-tts.md)