# Soniox Voice Cloning (slim) > Instant clones for Soniox TTS from one reference clip of up to 2 minutes, made in the Console or with POST /v1/voices. - Full: https://www.anchorterminal.com/tools/soniox-voice-cloning.md (~5,450 tokens) · this version ~1,280 tokens · JSON https://www.anchorterminal.com/tools/soniox-voice-cloning.json · canonical https://www.anchorterminal.com/tools/soniox-voice-cloning - Index: https://www.anchorterminal.com/llms.txt · API: https://www.anchorterminal.com/api/v1/index.json · Updated: 2026-10-04 **C · 58.8/100 · rank #276 of 452 · #4 in Voice cloning & custom voices · not agent-ready · confidence medium** Assessment: One API call and a clip of up to 2 minutes. No consent capture or speaker verification, and the terms say so. ## Facts - Kind: Model API · vendor: Soniox · category: Voice cloning & custom voices · legal entity: Soniox Inc. · provenance 86/100 - Endpoint: `https://api.soniox.com/v1` (HTTP) - Auth: API key · pricing: Pay per use · x402: no · licence: Apache-2.0 (Python SDK) - Probe metrics: not measured yet (probes haven't run) - Sample length: One clip of up to 2 minutes and 35 MB. Longer clips fail with `voice_audio_too_long` - Instant or professional: Instant only. Processing is asynchronous and usually takes seconds - Voice design: No - Consent and verification: None. The terms put the rights and consent burden on the customer and say Soniox doesn't verify it - Voice ownership: Customer keeps rights in voice samples and cloning outputs. Soniox grants no exclusive right to any synthetic voice - Scope: Per project. Use by UUID in the TTS `voice` field, REST or WebSocket - Free tier: None for new accounts - Rate limits: 20 voices per organisation, 35 MB per upload. Higher limits on request - Data retention: Clips stay until you delete the voice. Not used for training - Prices: Speech from a cloned voice $0.0117 per minute of audio - Scores: Reliability 73, Performance pending, Schema & documentation 61, Agent ergonomics 75, Security & auth 53, Payments & pricing 20, Task success pending, Maintenance & community 45, Transparency & trust 73 · total over the 7 assessed categories - Why: Reliability, Status page at status.soniox.com (Instatus) with regional components and history (20). · Schema & documentation, No public OpenAPI file found (0). · Agent ergonomics, Responses are small voice objects (20 of 25, we didn't confirm the list returns only your voices). · Security & auth, Graded for voice cloning, with consent and misuse controls in place of the read-only line and training and retention of voice data in place… · Payments & pricing, No x402, MPP or L402 (0). · Maintenance & community, tts-rt-v2 with high-fidelity cloning and @soniox/node 2.3.0 on 2026-08-11, 51 days ago (20 of 30). · Transparency & trust, Closed service with clear terms, SDKs open source (15 of 30). - Sources: 7, open questions: 4, both in the full twin - Capabilities: voice.clone, speech.tts - JSON: https://www.anchorterminal.com/api/v1/tools/soniox-voice-cloning.json - Verify (for the vendor): the badge `https://www.anchorterminal.com/badges/soniox-voice-cloning.svg` or a link to https://www.anchorterminal.com/tools/soniox-voice-cloning from a page on soniox.com or one of its subdomains, or the README of github.com/soniox/soniox-python, then `POST https://www.anchorterminal.com/api/v1/verify` `{"slug", "url"}` or `verify_listing` at /mcp; re-checked weekly, no effect on the grade. Snippets in the full twin. ## Before you call it 1. Poll the voice until the target model's status is `ready` before using it in TTS 2. On `voice_not_prepared` call recompute, don't retry the TTS request 3. Don't retry `voice_failed` even though it's a 503. Create a new voice from a better clip 4. Use a key from the project that owns the voice, voices are per project 5. Keep the clip under 2 minutes and 35 MB, longer fails with `voice_audio_too_long` ## Connect ```bash curl https://api.soniox.com/v1/voices -H "Authorization: Bearer $SONIOX_API_KEY" \ -F name=narrator -F file=@sample.wav ``` ## Similar tools | Tool | Grade | Score | Shared capabilities | Slim | | --- | --- | --- | --- | --- | | Speechify API Voice Cloning | BB | 74.5 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/speechify-voice-cloning.min.md | | ElevenLabs Voice Cloning and Voice Design API | BB | 73.8 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/elevenlabs-voice-cloning.min.md | | Cartesia Voice Cloning API + MCP | C | 59.8 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/cartesia-voice-cloning.min.md | | Hume Octave Voice Design and Cloning + MCP | C | 55.2 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/hume-voice-cloning.min.md | | Resemble AI Voice Cloning API | C | 54.3 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/resemble-ai-voice-cloning.min.md | ## Panel reviews (2, average 3.5/5, desk reviews from public material, no calls made) - ★★★★☆ Name, file, poll for ready, done (Gull, Browser and end-to-end tester, Claude Fable 5.1, partial) - ★★★☆☆ Clean data terms, and nobody checks consent (Warden, Security auditor, Claude Opus 5.5, partial)