# MiniMax Voice Clone and Voice Design (slim) > MiniMax's API clones a voice from a 10 second to 5 minute audio sample and designs voices from a text description. The resulting voice IDs work with its speech-2.8 text-to-speech models. - Full: https://www.anchorterminal.com/tools/minimax-voice-cloning.md (~6,850 tokens) · this version ~1,680 tokens · JSON https://www.anchorterminal.com/tools/minimax-voice-cloning.json · canonical https://www.anchorterminal.com/tools/minimax-voice-cloning - Index: https://www.anchorterminal.com/llms.txt · API: https://www.anchorterminal.com/api/v1/index.json · Updated: 2026-10-09 **D · 53.1/100 · rank #640 of 842 · #7 in Voice cloning & custom voices · not agent-ready · confidence medium** Assessment: Typed OpenAPI files, published prices of $1.50 a cloned voice and $3 a designed voice, and optional transcript and watermark checks. Cloning needs individual or enterprise account verification, the terms and privacy policy could not be read without a browser, and no statement on training with uploaded voice audio was found in the docs. ## Facts - Kind: Model API · vendor: MiniMax · category: Voice cloning & custom voices · legal entity: Nanonoble Pte. Ltd. · provenance 80/100 - Endpoint: `https://api.minimax.io` (HTTP, stdio) - Auth: API key · pricing: Pay per use · x402: no · licence: Proprietary API under MiniMax's platform terms. The two official MCP servers are MIT - Probe metrics: not measured yet (probes haven't run) - Sample length: One file of 10 seconds to 5 minutes, mp3, m4a or wav, up to 20 MB. An optional prompt clip under 8 seconds with a matching transcript improves similarity - Instant vs professional: Rapid cloning only, returned in one synchronous call. No slower professional tier in the API - Voice design: `POST /v1/voice_design` takes a description and a preview text of up to 500 characters and returns a `voice_id` with hex-encoded preview audio - Consent and verification: Individual or enterprise verification of the account is required. No speaker consent field. Optional `text_validation` transcript check and optional `aigc_watermark` tone on the preview - Voice ownership: Not established. The terms of service did not render for our reader. A voice can be used only by the account that created it (error 2042) - Languages: Any language for the sample, 40 languages for synthesis - Endpoints: `POST /v1/files/upload`, `POST /v1/voice_clone`, `POST /v1/voice_design`, `POST /v1/get_voice`, `POST /v1/delete_voice` - Free tier: None published for cloning or voice design - Rate limits: Voice cloning 60 requests a minute, voice design 20, T2A 60. Audio subscriptions set 10 to 800 requests a minute - Data retention: An unused voice is deleted after 7 days and a used one is kept until deleted through the API. No retention period for uploaded audio was found - MCP server: Official, stdio or SSE, `minimax-mcp` 0.0.19 on PyPI and `minimax-mcp-js` 0.0.18 on npm, 8 tools, 3 for voices - Prices: Rapid voice clone $1.50 per transaction; Voice design $3 per transaction; Speech from a clone, speech-2.8-hd $100 per 1M characters; Speech from a clone, speech-2.8-turbo $60 per 1M characters - Scores: Reliability 70, Performance pending, Schema & documentation 78, Agent ergonomics 56, Security & auth 39, Payments & pricing 20, Task success pending, Maintenance & community 38, Transparency & trust 53 · total over the 7 assessed categories - Why: Reliability, Atlassian Statuspage at status.minimax.io with a Text-to-Speech component and no separate component for cloning (20). · Schema & documentation, OpenAPI 3.1 files for cloning, upload, voice design and voice management (25). · Agent ergonomics, The clone call returns a preview URL and billing counts, but voice design returns preview audio as hex inside the JSON body, and listing wit… · Security & auth, Graded for voice cloning, with consent and misuse controls in place of the read-only line and training and retention of voice data in place… · Payments & pricing, No x402, MPP or L402 (0). · Maintenance & community, The newest dated change that touches this product is `minimax-mcp` 0.0.19 and `minimax-mcp-js` 0.0.18 on 21 August 2026, 48 days ago. · Transparency & trust, Closed API under platform terms that exist but did not render for our reader, with MIT MCP servers (12 of 30). - Sources: 24, open questions: 8, both in the full twin - Capabilities: voice.clone, voice.design, speech.tts - JSON: https://www.anchorterminal.com/api/v1/tools/minimax-voice-cloning.json - Verify (for the vendor): the badge `https://www.anchorterminal.com/badges/minimax-voice-cloning.svg` or a link to https://www.anchorterminal.com/tools/minimax-voice-cloning from a page on minimax.io or one of its subdomains, or the README of github.com/MiniMax-AI/MiniMax-MCP, then `POST https://www.anchorterminal.com/api/v1/verify` `{"slug", "url"}` or `verify_listing` at /mcp; re-checked weekly, no effect on the grade. Snippets in the full twin. ## Before you call it 1. Upload the sample to `POST /v1/files/upload` with `purpose=voice_clone`, then pass the integer `file_id` and a new `voice_id` of 8 to 256 characters to `POST /v1/voice_clone` 2. Synthesise once with T2A within 7 days or the voice is deleted. A preview inside the clone call does not count 3. Read `base_resp.status_code` on every response. Errors such as 2038 (no cloning permission) and 2039 (duplicate `voice_id`) arrive in the body 4. Voice design returns `trial_audio` as hex in the JSON body. Write it to a file and keep it out of the model context 5. Send `voice_type` to `POST /v1/get_voice`, because `all` also returns the 300+ system voices ## Connect ```bash curl --location 'https://api.minimax.io/v1/voice_clone' \ --header "Authorization: Bearer $MINIMAX_API_KEY" \ --header 'Content-Type: application/json' \ --data '{"file_id": 123456789, "voice_id": "MiniMax001"}' ``` ## Similar tools | Tool | Grade | Score | Shared capabilities | Slim | | --- | --- | --- | --- | --- | | ElevenLabs Voice Cloning and Voice Design API | BB | 73.6 | voice.clone, voice.design, speech.tts | https://www.anchorterminal.com/tools/elevenlabs-voice-cloning.min.md | | Hume Octave Voice Design and Cloning + MCP | C | 54.6 | voice.clone, voice.design, speech.tts | https://www.anchorterminal.com/tools/hume-voice-cloning.min.md | | Fish Audio Voice Cloning API | D | 51.3 | voice.clone, voice.design, speech.tts | https://www.anchorterminal.com/tools/fish-audio-voice-cloning.min.md | | Speechify API Voice Cloning | BB | 74.1 | voice.clone, speech.tts | https://www.anchorterminal.com/tools/speechify-voice-cloning.min.md | | Fish Audio TTS API | C | 60.7 | speech.tts, voice.clone | https://www.anchorterminal.com/tools/fish-audio-tts.min.md | ## Panel reviews (0, desk reviews from public material, no calls made)