# Mistral Voxtral Transcribe (slim) > Mistral AI's speech-to-text API. Voxtral Mini Transcribe 2 transcribes files of up to about three hours with diarisation, word timestamps and context biasing, and Voxtral Realtime transcribes live audio over a WebSocket. - Full: https://www.anchorterminal.com/tools/mistral-voxtral-transcribe.md (~7,650 tokens) · this version ~1,530 tokens · JSON https://www.anchorterminal.com/tools/mistral-voxtral-transcribe.json · canonical https://www.anchorterminal.com/tools/mistral-voxtral-transcribe - Index: https://www.anchorterminal.com/llms.txt · API: https://www.anchorterminal.com/api/v1/index.json · Updated: 2026-10-09 **B · 64.1/100 · rank #329 of 842 · #10 in Speech-to-text · not agent-ready · confidence medium** Assessment: Batch transcription costs $0.003 a minute and takes one multipart call with a file, a URL or an uploaded file ID. Rate limit numbers are shown only in the console, and timestamps cannot be combined with a set language, nor diarisation with the realtime model. ## Facts - Kind: Model API · vendor: Mistral AI · category: Speech-to-text · legal entity: Mistral AI (RCS Paris 952 418 325) · provenance 91/100 - Endpoint: `https://api.mistral.ai/v1` (HTTP, websocket) - Auth: API key · pricing: Freemium · x402: no · licence: Proprietary hosted service under Mistral's commercial terms. Voxtral Mini 4B Realtime weights and the SDKs are Apache-2.0 - Probe metrics: not measured yet (probes haven't run) - Models: `voxtral-mini-latest` (Voxtral Mini Transcribe 2, `voxtral-mini-2602`) for files, `voxtral-mini-transcribe-realtime-2602` for live audio - Languages: 13: English, Chinese, Hindi, Spanish, Arabic, French, Portuguese, Russian, German, Japanese, Korean, Italian and Dutch, with language detection - Max audio: About 3 hours a request on Voxtral Mini Transcribe 2 - Diarisation: `diarize` on the batch endpoint only - Extras: Segment or word timestamps, `context_bias` of up to 100 terms (tuned for English), SSE streaming of a file transcription - Streaming latency: Vendor says configurable below 200 ms through `target_streaming_delay_ms` - Realtime input: PCM audio such as `pcm_s16le` at 16 kHz over a WebSocket - Rate limits: Audio seconds a minute and a month per workspace, numbers shown in the Admin Panel - Free tier: Free mode, no card, console limits - Data retention: 30 rolling days for abuse monitoring. Zero data retention on request for paid plans covers `/v1/audio/transcriptions` - Regions: Global endpoint, with opt-in EU and US endpoints at 1.1 times list price - Open weights: Voxtral Mini 4B Realtime 2602 on Hugging Face, Apache-2.0 - MCP server: None for transcription - Prices: Voxtral Mini Transcribe 2 (batch) $0.003 per minute of audio; Voxtral Mini Transcribe Realtime $0.006 per minute of audio - Scores: Reliability 53, Performance pending, Schema & documentation 85, Agent ergonomics 77, Security & auth 70, Payments & pricing 40, Task success pending, Maintenance & community 66, Transparency & trust 82 · negative events -3 · total over the 7 assessed categories - Why: Reliability, Hosted reading. · Schema & documentation, Model reading. · Agent ergonomics, API reading, as for the other speech-to-text listings. · Security & auth, Model reading, scored like the other Mistral listings. · Payments & pricing, No x402, MPP or L402 (0). · Maintenance & community, Model reading. · Transparency & trust, The hosted service is closed under commercial terms. - Sources: 23, open questions: 8, both in the full twin - Capabilities: speech.stt, speech.batch, speech.streaming, speech.diarisation, speech.languages - JSON: https://www.anchorterminal.com/api/v1/tools/mistral-voxtral-transcribe.json - Verify (for the vendor): the badge `https://www.anchorterminal.com/badges/mistral-voxtral-transcribe.svg` or a link to https://www.anchorterminal.com/tools/mistral-voxtral-transcribe from a page on mistral.ai or one of its subdomains, or the README of github.com/mistralai/client-python, then `POST https://www.anchorterminal.com/api/v1/verify` `{"slug", "url"}` or `verify_listing` at /mcp; re-checked weekly, no effect on the grade. Snippets in the full twin. ## Before you call it 1. Send `model=voxtral-mini-latest` and one of `file`, `file_url` or `file_id` as multipart form fields to `/v1/audio/transcriptions` 2. Leave `language` out when you ask for `timestamp_granularities`, the docs say the two are not compatible 3. Use the batch endpoint for diarisation. `voxtral-mini-transcribe-realtime-2602` does not accept `diarize` 4. Pin `voxtral-mini-2602` if output must not change, since `-latest` aliases can move 5. Mint browser tokens with `POST /v1/client/sessions` close to connection time and pass them in `Sec-WebSocket-Protocol`, never the API key ## Connect ```bash pip install mistralai # realtime: pip install "mistralai[realtime]" # or: npm i @mistralai/mistralai ``` ```bash curl --location 'https://api.mistral.ai/v1/audio/transcriptions' \ --header "x-api-key: $MISTRAL_API_KEY" \ --form 'file_url="https://docs.mistral.ai/audio/obama.mp3"' \ --form 'model="voxtral-mini-latest"' ``` ## Similar tools | Tool | Grade | Score | Shared capabilities | Slim | | --- | --- | --- | --- | --- | | Amazon Transcribe | BB | 73.4 | speech.stt, speech.streaming, speech.batch, speech.diarisation, speech.languages | https://www.anchorterminal.com/tools/amazon-transcribe.min.md | | Azure AI Speech speech-to-text | BB | 73 | speech.stt, speech.streaming, speech.batch, speech.diarisation, speech.languages | https://www.anchorterminal.com/tools/azure-speech-to-text.min.md | | Deepgram Speech-to-Text (Nova-3, Flux) | BB | 70.3 | speech.stt, speech.streaming, speech.batch, speech.diarisation, speech.languages | https://www.anchorterminal.com/tools/deepgram-stt.min.md | | Google Cloud Speech-to-Text | BB | 70.2 | speech.stt, speech.streaming, speech.batch, speech.diarisation, speech.languages | https://www.anchorterminal.com/tools/google-speech-to-text.min.md | | Gladia Speech-to-Text API + MCP | B | 69.5 | speech.stt, speech.streaming, speech.batch, speech.diarisation, speech.languages | https://www.anchorterminal.com/tools/gladia-stt.min.md | ## Panel reviews (0, desk reviews from public material, no calls made)