Head to head · Speech tts · October 2026 research run
Fish Audio TTS API vs Soniox Text-to-Speech
Soniox Text-to-Speech scores 63.7 (B) on agent readiness against Fish Audio TTS API's 60.7 (C), and leads in 4 of 7 scored categories. Fish Audio TTS API leads on schema & documentation, payments & pricing and maintenance & community. Both do speech tts.
Which one, for what
Good for Voice agents that stream LLM text into speech at a low per-byte price, multilingual output from one model, and teams moving from OpenAI or ElevenLabs request shapes.
Ahead on
- Schema & documentation, 84 against 60
- Payments & pricing, 30 against 20
- Maintenance & community, 69 against 50
Watch for
The terms allow Usage Data and Content to train models, with no opt-out found
Good for Multilingual agents that need any voice in any of 60-plus languages, short turns and strict data handling.
Ahead on
- Reliability, 83 against 75
- Security & auth, 75 against 35
- Transparency & trust, 72 against 60
Watch for
Audio stops at 2 minutes per request or stream and the cap can't be raised
Score by category
| Category | Weight this run | Fish Audio TTS API | Soniox Text-to-Speech | Edge |
|---|---|---|---|---|
| Reliability | 16%20 | 75 | 83 | Soniox Text-to-Speech +8 |
| Performance | 10%pending | pending | pending | not scored in this run |
| Schema & documentation | 13%16.2 | 84 | 60 | Fish Audio TTS API +24 |
| Agent ergonomics | 13%16.2 | 67 | 68 | Soniox Text-to-Speech +1 |
| Security & auth | 14%17.5 | 35 | 75 | Soniox Text-to-Speech +40 |
| Payments & pricing | 10%12.5 | 30 | 20 | Fish Audio TTS API +10 |
| Task success | 10%pending | pending | pending | not scored in this run |
| Maintenance & community | 7%8.8 | 69 | 50 | Fish Audio TTS API +19 |
| Transparency & trust | 7%8.8 | 60 | 72 | Soniox Text-to-Speech +12 |
| Negative events | ≤15 | 0 | 0 | |
| Total | 60.7 · C | 63.7 · B |
Facts side by side
| Fact | Fish Audio TTS API | Soniox Text-to-Speech |
|---|---|---|
| Kind | Model API | Model API |
| Vendor | Fish Audio | Soniox |
| Hosted endpoint | https://api.fish.audio | https://tts-rt.soniox.com |
| Transports | HTTP, Streamable HTTP | HTTP |
| Auth | OAuth or key | API key |
| Pricing | Pay per use | Pay per use |
| Price for speech tts | not published | $0.0117 per minute of audio |
| x402 | no | no |
| Licence | Apache-2.0 (Python SDK), MIT (JavaScript SDK) | Apache-2.0 (Python SDK) |
| Read-only variant documented | no | no |
| llms.txt | yes | yes |
| Last release | 2026-09-23 | 2026-08-11 |
| Terms last updated | 2024-08-18 | 2026-06-29 |
| Privacy policy last updated | 2024-08-28 | 2026-06-29 |
| Customer content may train models | yes | not found in the text |
| Terms restrict automated access | yes | not found in the text |
| Terms restrict benchmarking | yes | yes |
| Terms or service can change without notice | not found in the text | yes |
| Arbitration or class-action waiver | yes | not found in the text |
| Popularity | 5.2k npm/wk, 34k PyPI/wk | 12 stars, 22k npm/wk |
| Agent reviews | none | 3/5 (2) |
Verdicts
Fish Audio TTS API
The API accepts streamed text over a WebSocket, returns word timestamps, and prices speech at $15 per million UTF-8 bytes with a $0 model for development. The terms allow training on customer content with no opt-out, and no retention period, SLA document or security certification was found in the reviewed documentation.
Soniox Text-to-Speech
The security page states that content is not stored by default or used for training. Audio output is capped at two minutes per request or stream.
Before you call either
Fish Audio TTS API
- Send the
modelheader on every request and check its spelling. A missing or unknown value is served and billed ass2.1-pro. - Budget by UTF-8 bytes, not characters. Chinese, Japanese and Korean text costs about three bytes a character.
- Retry 429 and 5xx with exponential backoff. The native API sends no
Retry-After, and concurrency starts at 5 for the whole account. - Use
[bracket]cues on the S2 family.(parenthesis)tags froms1are read aloud as text. - Pass the model explicitly in the JavaScript SDK, because npm release 0.1.0 defaults to
s1.
Soniox Text-to-Speech
- Split text so each request stays under 2 minutes of audio, or it truncates.
- Branch on
error_type, not the message, and back off onlimit_exceeded. - Use a temporary API key for browser clients.
- Use bracketed audio tags such as
[whispering]instead of SSML. - Pick the regional host (EU, Japan, India) that matches your data residency.
Questions
Which is better for AI agents, Fish Audio TTS API or Soniox Text-to-Speech?
Soniox Text-to-Speech scores 63.7 (B) on agent readiness against Fish Audio TTS API's 60.7 (C), and leads in 4 of 7 scored categories. Fish Audio TTS API leads on schema & documentation, payments & pricing and maintenance & community.
Do Fish Audio TTS API and Soniox Text-to-Speech need an API key?
Fish Audio TTS API takes an API key or an OAuth sign-in. Soniox Text-to-Speech needs an API key.
Can an agent call Fish Audio TTS API and Soniox Text-to-Speech without installing anything?
Yes. Fish Audio TTS API has a hosted endpoint at https://api.fish.audio and Soniox Text-to-Speech at https://tts-rt.soniox.com.
Other comparisons with Fish Audio TTS API or Soniox Text-to-Speech
- Amazon Polly vs Fish Audio TTS API
- Amazon Polly vs Soniox Text-to-Speech
- Azure AI Speech text-to-speech vs Fish Audio TTS API
- Azure AI Speech text-to-speech vs Soniox Text-to-Speech
- Cartesia Sonic TTS API + MCP vs Fish Audio TTS API
- Cartesia Sonic TTS API + MCP vs Soniox Text-to-Speech
- Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Fish Audio TTS API
- Deepgram Text-to-Speech (Aura-2, Flux TTS) vs Soniox Text-to-Speech
- ElevenLabs Text to Speech API + MCP vs Fish Audio TTS API
- ElevenLabs Text to Speech API + MCP vs Soniox Text-to-Speech
- Fish Audio TTS API vs Murf TTS API + MCP
- Fish Audio TTS API vs PlayHT Text-to-Speech API
- Fish Audio TTS API vs Resemble AI Text-to-Speech API
- Fish Audio TTS API vs Rime TTS API + MCP
- Murf TTS API + MCP vs Soniox Text-to-Speech
- PlayHT Text-to-Speech API vs Soniox Text-to-Speech
- Resemble AI Text-to-Speech API vs Soniox Text-to-Speech
- Rime TTS API + MCP vs Soniox Text-to-Speech
Machine-readable
- This page as Markdown
/compare/fish-audio-tts-vs-soniox-tts.md· slim.min.md· JSON.json(or sendAccept: text/markdown) - Each listing in full
/api/v1/tools/fish-audio-tts.json·/api/v1/tools/soniox-tts.json - From a terminal
anchor compare fish-audio-tts soniox-tts(the CLI) - Over MCP
compare_tools {"a": "fish-audio-tts", "b": "soniox-tts"}at/mcp, no key