# Voice cloning and custom voice APIs > 10 voice cloning & custom voices ranked by the Anchor benchmark. Leader Speechify API Voice Cloning (BB). APIs that make a new voice from a sample, or design one from a description. Compared on how much audio a clone needs, how close it sounds, consent and verification checks, who owns the voice, and price. - Canonical: https://www.anchorterminal.com/categories/voice-cloning - Markdown: https://www.anchorterminal.com/categories/voice-cloning.md (~2,600 tokens) - Slim: https://www.anchorterminal.com/categories/voice-cloning.min.md (~580 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/categories/voice-cloning.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-04 APIs that make a new voice from a sample, or design one from a description. Compared on how much audio a clone needs, how close it sounds, consent and verification checks, who owns the voice, and price. - Tools ranked: 10 · agent-ready (BB or better): 2 · accept x402: 0 · hosted endpoints: 10 · desk reviews by the panel: 26 - JSON: https://www.anchorterminal.com/api/v1/tools.json (list) · https://www.anchorterminal.com/api/v1/rankings.json (ranked) · https://www.anchorterminal.com/api/v1/x402.json (payable) · https://www.anchorterminal.com/api/v1/capabilities.json (by capability) - Grades run AA, A, BB, B, C, D, E, F · methodology: https://www.anchorterminal.com/benchmark/ - Capabilities in this category: voice.clone, voice.design, voice.consent, speech.tts - https://letme.dev/voice.clone picks the top-graded tool in this list and says how to call it direct; calling through letme comes later (https://www.anchorterminal.com/letme/index.md) ## Ranking | # | Tool | Vendor | Kind | Category | Grade | Score | Confidence | x402 | Auth | Where | Reviews | Page | | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | | 47 | Speechify API Voice Cloning | Speechify | Model API | Cloning | BB | 74.5 | medium | no | API key | hosted | 3.6/5 (8) | https://www.anchorterminal.com/tools/speechify-voice-cloning.md | | 53 | ElevenLabs Voice Cloning and Voice Design API | ElevenLabs | Model API | Cloning | BB | 73.8 | high | no | API key | hosted + local | 3.5/5 (2) | https://www.anchorterminal.com/tools/elevenlabs-voice-cloning.md | | 257 | Cartesia Voice Cloning API + MCP | Cartesia | Model API | Cloning | C | 59.8 | medium | no | OAuth or key | hosted + local | 2.5/5 (2) | https://www.anchorterminal.com/tools/cartesia-voice-cloning.md | | 276 | Soniox Voice Cloning | Soniox | Model API | Cloning | C | 58.8 | medium | no | API key | hosted | 3.5/5 (2) | https://www.anchorterminal.com/tools/soniox-voice-cloning.md | | 316 | Hume Octave Voice Design and Cloning + MCP | Hume AI | Model API | Cloning | C | 55.2 | medium | no | API key | hosted + local | 2/5 (2) | https://www.anchorterminal.com/tools/hume-voice-cloning.md | | 325 | Resemble AI Voice Cloning API | Resemble AI | Model API | Cloning | C | 54.3 | medium | no | API key | hosted | 2.5/5 (2) | https://www.anchorterminal.com/tools/resemble-ai-voice-cloning.md | | 348 | Fish Audio Voice Cloning API | Fish Audio | Model API | Cloning | D | 51.5 | medium | no | API key | hosted | 2.5/5 (2) | https://www.anchorterminal.com/tools/fish-audio-voice-cloning.md | | 394 | Murf Voice Cloning API | Murf | Model API | Cloning | D | 46.5 | medium | no | API key | hosted | 2/5 (2) | https://www.anchorterminal.com/tools/murf-voice-cloning.md | | 420 | Ultravox Voice Cloning | Ultravox (Fixie.ai) | Model API | Cloning | E | 41 | medium | no | API key | hosted | 2.5/5 (2) | https://www.anchorterminal.com/tools/ultravox-voice-cloning.md | | not ranked, shut down | PlayHT Voice Cloning API | PlayHT | Model API | Cloning | F | 4.2 | medium | no | API key | hosted | 1/5 (2) | https://www.anchorterminal.com/tools/playht-voice-cloning.md | Scores are from public evidence against the published checklist (https://www.anchorterminal.com/benchmark/), with Performance and Task success pending. p95 latency and context cost come from our probes, which haven't run yet. ## Summaries ### 47. Speechify API Voice Cloning, BB (74.5) Zero-shot voice cloning on the SpeechifyAI API from a 10 to 30 second sample, with verified spoken consent. Verified consent on every clone, phrase and speaker matched, recording kept. No cloning on the Free plan. - Page: https://www.anchorterminal.com/tools/speechify-voice-cloning · Markdown: https://www.anchorterminal.com/tools/speechify-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/speechify-voice-cloning.json - Capabilities: voice.clone, voice.consent, speech.tts · endpoint: `https://api.speechify.ai/v1` ### 53. ElevenLabs Voice Cloning and Voice Design API, BB (73.8) ElevenLabs' voice cloning service creates voices from audio samples. Professional cloning uses verified recordings, while Voice Design generates voices from text prompts. Instant clones from 1 to 2 minutes of audio on the $6 Starter plan. Instant clones rely on an attestation, not a check. - Page: https://www.anchorterminal.com/tools/elevenlabs-voice-cloning · Markdown: https://www.anchorterminal.com/tools/elevenlabs-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/elevenlabs-voice-cloning.json - Capabilities: voice.clone, voice.design, voice.consent, speech.tts · endpoint: `https://api.elevenlabs.io/v1` ### 257. Cartesia Voice Cloning API + MCP, C (59.8) Instant voice clones for Sonic from 10 to 60 seconds of audio, and Pro Voice Clones fine-tuned on 30 minutes or more, trained in up to 3 hours. Instant clones from 10 seconds of audio, up to 60 seconds used on Sonic 3.6. No consent or speaker verification in the clone API. - Page: https://www.anchorterminal.com/tools/cartesia-voice-cloning · Markdown: https://www.anchorterminal.com/tools/cartesia-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/cartesia-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://api.cartesia.ai` ### 276. Soniox Voice Cloning, C (58.8) Instant clones for Soniox TTS from one reference clip of up to 2 minutes, made in the Console or with `POST /v1/voices`. One API call and a clip of up to 2 minutes. No consent capture or speaker verification, and the terms say so. - Page: https://www.anchorterminal.com/tools/soniox-voice-cloning · Markdown: https://www.anchorterminal.com/tools/soniox-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/soniox-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://api.soniox.com/v1` ### 316. Hume Octave Voice Design and Cloning + MCP, C (55.2) Custom voices for Hume's Octave TTS and EVI. Voice design from a natural-language description, saved with one call. Cloning over the API is Enterprise-only. - Page: https://www.anchorterminal.com/tools/hume-voice-cloning · Markdown: https://www.anchorterminal.com/tools/hume-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/hume-voice-cloning.json - Capabilities: voice.clone, voice.design, speech.tts · endpoint: `https://api.hume.ai/v0/tts` ### 325. Resemble AI Voice Cloning API, C (54.3) Resemble AI's service for creating custom voices for speech generation. Rapid clones from 10 seconds of audio, ready in under a minute. Cloning API only on Business ($1,000 a month) and Enterprise. - Page: https://www.anchorterminal.com/tools/resemble-ai-voice-cloning · Markdown: https://www.anchorterminal.com/tools/resemble-ai-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/resemble-ai-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://app.resemble.ai/api/v2` ### 348. Fish Audio Voice Cloning API, D (51.5) Fish Audio's API creates reusable voice clones from audio samples and generates candidate voices from text prompts. Usable clone from about 10 seconds of audio, available as soon as it's created. No consent or speaker verification in the API. - Page: https://www.anchorterminal.com/tools/fish-audio-voice-cloning · Markdown: https://www.anchorterminal.com/tools/fish-audio-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/fish-audio-voice-cloning.json - Capabilities: voice.clone, voice.design, speech.tts · endpoint: `https://api.fish.audio` ### 394. Murf Voice Cloning API, D (46.5) Murf's service for creating voices from audio samples, with instant cloning and a managed professional option. One sample of up to 30 seconds, ready in a few minutes. Enterprise only, with no self-serve or trial access. - Page: https://www.anchorterminal.com/tools/murf-voice-cloning · Markdown: https://www.anchorterminal.com/tools/murf-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/murf-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://api.murf.ai/v1` ### 420. Ultravox Voice Cloning, E (41) Instant voice cloning inside Ultravox Realtime. One multipart request with a 30 to 60 second sample. No consent or speaker verification, only a warranty in the terms. - Page: https://www.anchorterminal.com/tools/ultravox-voice-cloning · Markdown: https://www.anchorterminal.com/tools/ultravox-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/ultravox-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://api.ultravox.ai/api` ### PlayHT Voice Cloning API, F (4.2), retired, not ranked Discontinued PlayHT service for creating voices from uploaded audio samples. The API reference is still readable at docs.play.ht, useful for mapping old integrations. Shut down, api.play.ht doesn't resolve. - Page: https://www.anchorterminal.com/tools/playht-voice-cloning · Markdown: https://www.anchorterminal.com/tools/playht-voice-cloning.md · JSON: https://www.anchorterminal.com/api/v1/tools/playht-voice-cloning.json - Capabilities: voice.clone, speech.tts · endpoint: `https://api.play.ht/api/v2` ## How we test this category The same consented sample through every API, short and long. We judge similarity blind, check what consent or verification each vendor asks for, note the licence on the resulting voice, and price a clone and an hour of speech from it. This test hasn't run yet, so Task success is pending and the grades here come from the categories assessed from public evidence.