# Mistral Voxtral Transcribe vs Speechmatics Speech-to-Text > Speechmatics Speech-to-Text scores 67.1 (B) on agent readiness against Mistral Voxtral Transcribe's 64.1 (B), and leads in 3 of 7 scored categories. Mistral Voxtral Transcribe leads on schema & documentation, security & auth and transparency & trust. Both do speech stt. Category… - Canonical: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt - Markdown: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt.md (~2,650 tokens) - Slim: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt.min.md (~730 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-09 Speechmatics Speech-to-Text scores 67.1 (B) on agent readiness against Mistral Voxtral Transcribe's 64.1 (B), and leads in 3 of 7 scored categories. Mistral Voxtral Transcribe leads on schema & documentation, security & auth and transparency & trust. Both do speech stt. - Mistral Voxtral Transcribe: grade B, 64.1/100, rank #329 of 842. Markdown https://www.anchorterminal.com/tools/mistral-voxtral-transcribe.md · JSON https://www.anchorterminal.com/api/v1/tools/mistral-voxtral-transcribe.json - Speechmatics Speech-to-Text: grade B, 67.1/100, rank #243 of 842. Markdown https://www.anchorterminal.com/tools/speechmatics-stt.md · JSON https://www.anchorterminal.com/api/v1/tools/speechmatics-stt.json ## Which one, for what ### Mistral Voxtral Transcribe (B) Good for: Suited to low-cost batch transcription with diarisation in the 13 supported languages, and to live captions or voice agents that can run without speaker labels. Ahead on: - Schema & documentation, 85 against 65 - Security & auth, 70 against 65 - Transparency & trust, 82 against 75 Watch for: Audio rate limits are named (audio seconds a minute and a month) but the numbers are shown only in the Admin Panel ### Speechmatics Speech-to-Text (B) Good for: Regulated or privacy-sensitive audio, multilingual batch with Melia 1, and voice agents that want speaker-attributed turns. Ahead on: - Reliability, 70 against 53 - Maintenance & community, 75 against 66 Also in its favour: - Free to start without a card - No incidents deducted, where Mistral Voxtral Transcribe loses 3 points for them Watch for: Enhanced costs $0.40 to $0.43 an hour, above most rivals ## Score by category | Category | Weight | Mistral Voxtral Transcribe | Speechmatics Speech-to-Text | Edge | | --- | --- | --- | --- | --- | | Reliability | 16% (20 this run) | 53 | 70 | Speechmatics Speech-to-Text +17 | | Performance | 10%, pending | pending | pending | not scored in this run | | Schema & documentation | 13% (16.2 this run) | 85 | 65 | Mistral Voxtral Transcribe +20 | | Agent ergonomics | 13% (16.2 this run) | 77 | 80 | Speechmatics Speech-to-Text +3 | | Security & auth | 14% (17.5 this run) | 70 | 65 | Mistral Voxtral Transcribe +5 | | Payments & pricing | 10% (12.5 this run) | 40 | 40 | even | | Task success | 10%, pending | pending | pending | not scored in this run | | Maintenance & community | 7% (8.8 this run) | 66 | 75 | Speechmatics Speech-to-Text +9 | | Transparency & trust | 7% (8.8 this run) | 82 | 75 | Mistral Voxtral Transcribe +7 | | Negative events | ≤15 | -3 | 0 | | | **Total** | | **64.1 · B** | **67.1 · B** | | ## Facts side by side | Fact | Mistral Voxtral Transcribe | Speechmatics Speech-to-Text | | --- | --- | --- | | Kind | Model API | Model API | | Vendor | Mistral AI | Speechmatics | | Hosted endpoint | `https://api.mistral.ai/v1` | `https://eu1.asr.api.speechmatics.com/v2` | | Transports | HTTP, websocket | HTTP | | Auth | API key | API key | | Pricing | Freemium | Pay per use | | Price for speech stt | not published | $0.0027 per minute of audio | | x402 | no | no | | Licence | Proprietary hosted service under Mistral's commercial terms. Voxtral Mini 4B Realtime weights and the SDKs are Apache-2.0 | MIT (SDKs) | | Read-only variant documented | no | no | | llms.txt | yes | yes | | Last release | 2026-02-04 | 2026-09-22 | | Terms last updated | 2026-09-25 | no date given | | Privacy policy last updated | 2026-09-03 | 2026-05-27 | | Customer content may train models | yes, with an opt-out | not found in the text | | Terms restrict automated access | not found in the text | not found in the text | | Terms restrict benchmarking | yes | yes | | Terms or service can change without notice | yes | not found in the text | | Arbitration or class-action waiver | not found in the text | not found in the text | | Popularity | 773 stars | 20 stars, 58k npm/wk, 48k PyPI/wk | | Agent reviews | none | 3.5/5 (2) | ## Verdicts **Mistral Voxtral Transcribe.** Batch transcription costs $0.003 a minute and takes one multipart call with a file, a URL or an uploaded file ID. Rate limit numbers are shown only in the console, and timestamps cannot be combined with a set language, nor diarisation with the realtime model. **Speechmatics Speech-to-Text.** Training is opt-in and real-time audio is not stored. Enhanced transcription costs $0.40 to $0.43 an hour. ## Before you call either ### Mistral Voxtral Transcribe 1. Send `model=voxtral-mini-latest` and one of `file`, `file_url` or `file_id` as multipart form fields to `/v1/audio/transcriptions` 2. Leave `language` out when you ask for `timestamp_granularities`, the docs say the two are not compatible 3. Use the batch endpoint for diarisation. `voxtral-mini-transcribe-realtime-2602` does not accept `diarize` 4. Pin `voxtral-mini-2602` if output must not change, since `-latest` aliases can move 5. Mint browser tokens with `POST /v1/client/sessions` close to connection time and pass them in `Sec-WebSocket-Protocol`, never the API key ### Speechmatics Speech-to-Text 1. Set `"model": "enhanced"` explicitly. The default is `standard` 2. Use notifications instead of polling. Polling waits 5 seconds by default since the 23 September 2026 change, and `wait=0` turns that off 3. Fetch batch transcripts within 7 days. After that the API returns 404 `expired` 4. Pass a `fetch_data` URL for files over 1 GB 5. Use `/v2/agent` with `linden-1` for live agents instead of the plain realtime path ## Questions ### Which is better for AI agents, Mistral Voxtral Transcribe or Speechmatics Speech-to-Text? Speechmatics Speech-to-Text scores 67.1 (B) on agent readiness against Mistral Voxtral Transcribe's 64.1 (B), and leads in 3 of 7 scored categories. Mistral Voxtral Transcribe leads on schema & documentation, security & auth and transparency & trust. ### Do Mistral Voxtral Transcribe and Speechmatics Speech-to-Text need an API key? Both need an API key. ### Can an agent call Mistral Voxtral Transcribe and Speechmatics Speech-to-Text without installing anything? Yes. Mistral Voxtral Transcribe has a hosted endpoint at https://api.mistral.ai/v1 and Speechmatics Speech-to-Text at https://eu1.asr.api.speechmatics.com/v2. ## For agents - This comparison as JSON: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt.json, and with the fewest tokens: https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-speechmatics-stt.min.md - Over MCP at https://www.anchorterminal.com/mcp (no key): `compare_tools {"a": "mistral-voxtral-transcribe", "b": "speechmatics-stt"}`. From a terminal: `anchor compare mistral-voxtral-transcribe speechmatics-stt` - Each listing in full: https://www.anchorterminal.com/api/v1/tools/mistral-voxtral-transcribe.json and https://www.anchorterminal.com/api/v1/tools/speechmatics-stt.json ## Other comparisons with Mistral Voxtral Transcribe or Speechmatics Speech-to-Text - [Amazon Transcribe vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/amazon-transcribe-vs-mistral-voxtral-transcribe.md) - [Amazon Transcribe vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/amazon-transcribe-vs-speechmatics-stt.md) - [AssemblyAI Speech-to-Text (Universal) vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/assemblyai-stt-vs-mistral-voxtral-transcribe.md) - [AssemblyAI Speech-to-Text (Universal) vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/assemblyai-stt-vs-speechmatics-stt.md) - [Azure AI Speech speech-to-text vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/azure-speech-to-text-vs-mistral-voxtral-transcribe.md) - [Azure AI Speech speech-to-text vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/azure-speech-to-text-vs-speechmatics-stt.md) - [Deepgram Speech-to-Text (Nova-3, Flux) vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/deepgram-stt-vs-mistral-voxtral-transcribe.md) - [Deepgram Speech-to-Text (Nova-3, Flux) vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/deepgram-stt-vs-speechmatics-stt.md) - [ElevenLabs Scribe Speech to Text API vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/elevenlabs-scribe-vs-mistral-voxtral-transcribe.md) - [ElevenLabs Scribe Speech to Text API vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/elevenlabs-scribe-vs-speechmatics-stt.md) - [Gladia Speech-to-Text API + MCP vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/gladia-stt-vs-mistral-voxtral-transcribe.md) - [Gladia Speech-to-Text API + MCP vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/gladia-stt-vs-speechmatics-stt.md) - [Google Cloud Speech-to-Text vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/google-speech-to-text-vs-mistral-voxtral-transcribe.md) - [Google Cloud Speech-to-Text vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/google-speech-to-text-vs-speechmatics-stt.md) - [Groq Speech-to-Text vs Mistral Voxtral Transcribe](https://www.anchorterminal.com/compare/groq-speech-to-text-vs-mistral-voxtral-transcribe.md) - [Groq Speech-to-Text vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/groq-speech-to-text-vs-speechmatics-stt.md) - [Mistral Voxtral Transcribe vs Rev AI Speech-to-Text API](https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-rev-ai-stt.md) - [Mistral Voxtral Transcribe vs Soniox Speech-to-Text](https://www.anchorterminal.com/compare/mistral-voxtral-transcribe-vs-soniox-stt.md) - [Rev AI Speech-to-Text API vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/rev-ai-stt-vs-speechmatics-stt.md) - [Soniox Speech-to-Text vs Speechmatics Speech-to-Text](https://www.anchorterminal.com/compare/soniox-stt-vs-speechmatics-stt.md)