# Mistral Voxtral Transcribe Review

> Mistral Voxtral Transcribe review: Low-cost EU-based transcription API with Apache-2.0 open-weight models. Verified pricing, features, and the strongest.

![Mistral Voxtral Transcribe logo](https://www.versusref.com/logos/voxtral.png)

Low-cost EU-based transcription API with Apache-2.0 open-weight models

Among the 43 speech-to-text tools we track, Mistral Voxtral Transcribe has the 37th-widest language coverage.

See pricing · facts verified Jul 20, 2026

## What we know about Mistral Voxtral Transcribe

Mistral Voxtral Transcribe sits in the speech-to-text apis category, where it is low-cost EU-based transcription API with Apache-2.0 open-weight models. We keep this Mistral Voxtral Transcribe profile grounded in primary sources, each fact dated to when we last confirmed it.

Mistral Voxtral Transcribe does not publish a public per-minute rate we have been able to verify, so the price figure here stays blank until we can confirm one. We would rather show nothing than a guessed number; if you have a current Mistral Voxtral Transcribe quote, it is the fastest way for us to close that gap.

On capabilities, Mistral Voxtral Transcribe covers language auto-detection, word-level timestamps, custom vocabulary / keyterm boosting, summarization endpoint, speech translation, and soc 2 type ii. Each of those is verified against Mistral Voxtral Transcribe's own docs or dashboard, not marketing copy.

For compliance, with Mistral Voxtral Transcribe: SOC 2 Type II is in place. If you are in a regulated space, confirm the current posture with Mistral Voxtral Transcribe before you commit, since these change plan by plan.

Among the 43 speech-to-text tools in our matrix, Mistral Voxtral Transcribe leads with the 37th-widest language coverage; the fact sheet below has the raw numbers behind that placement.

In total we track 23 verified facts for Mistral Voxtral Transcribe today, each linking the primary source it came from so you can check our work - and vendor claims we have not measured ourselves are labeled as such on the Mistral Voxtral Transcribe fact sheet below.

## Mistral Voxtral Transcribe pricing

Published rates: batch $0.003/min · streaming $0.006/min, verified Jul 20, 2026 ([source](https://mistral.ai/pricing/api)).

| Monthly volume | Batch bill | Streaming bill |
| --- | --- | --- |
| 1K min/mo | $3 | $6 |
| 10K min/mo | $30 | $60 |
| 100K min/mo | $300 | $600 |

Sticker rates only; diarization and PII redaction add-ons price in the stack builder.

## Fact sheet

### Pricing

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Batch price per audio minute | 0.003 $/audio-min | Jul 20 | [source](https://mistral.ai/pricing/api) |
| Streaming price per audio minute | 0.006 $/audio-min | Jul 20 | [source](https://mistral.ai/pricing/api) |
| Pricing model | usage | Jul 20 | [source](https://mistral.ai/pricing/api) |
| Published volume discounts | Batch API -50%, incl. /v1/audio/transcriptions | Jul 20 | [source](https://docs.mistral.ai/capabilities/batch) |

### Capabilities

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| WER (third-party benchmark) | 3.59 % WER | Jul 20 | [source](https://artificialanalysis.ai/speech-to-text) |
| WER (vendor-claimed) | ~4 % WER | Jul 20 | [source](https://mistral.ai/news/voxtral-transcribe-2/) |
| Streaming latency (vendor-claimed) | ~200 ms | Jul 20 | [source](https://mistral.ai/news/voxtral-transcribe-2/) |
| Languages supported | 13 | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text) |
| Language auto-detection | ✓  Yes | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text/offline_transcription) |
| Speaker diarization | ✓  Included | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text) |
| Word-level timestamps | ✓  Yes | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text/offline_transcription) |
| Custom vocabulary / keyterm boosting | ✓  Yes | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text/offline_transcription) |
| Summarization endpoint | ✓  Yes | Jul 20 | [source](https://mistral.ai/news/voxtral/) |
| Speech translation | ✓  Yes | Jul 20 | [source](https://mistral.ai/news/voxtral/) |
| Priced audio-intelligence add-ons | Voxtral Small audio understanding $0.004/min (chat API) | Jul 20 | [source](https://mistral.ai/pricing/api) |

### Compliance & trust

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| SOC 2 Type II | ✓  Yes | Jul 20 | [source](https://help.mistral.ai/en/articles/347638-do-you-have-soc-2-or-iso-27001-certification) |
| GDPR / EU data residency | ✓  Yes | Jul 20 | [source](https://help.mistral.ai/en/articles/347629-where-do-you-store-my-data-or-my-organization-s-data) |
| Self-host / on-prem option | ✓  Yes | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text) |
| Model weights license | Apache-2.0 | Jul 20 | [source](https://huggingface.co/mistralai/Voxtral-Small-24B-2507) |

### Build experience

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Official SDKs | Python, TypeScript | Jul 20 | [source](https://docs.mistral.ai/getting-started/clients) |
| Max file size / duration | 500 MB max file; up to 3 hours per request (V2) | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text) |
| Supported audio formats | WAV, MP3, FLAC, OGG, WEBM | Jul 20 | [source](https://docs.mistral.ai/resources/known-limitations) |
| Websocket streaming API | ✓  Yes | Jul 20 | [source](https://docs.mistral.ai/studio-api/audio/speech_to_text/realtime_transcription) |

## Mistral Voxtral Transcribe head-to-head

| Comparison | Record | Per use case |
| --- | --- | --- |
| [Mistral Voxtral Transcribe vs OpenAI Whisper (API)](https://www.versusref.com/stt/voxtral-vs-whisper/) | won 5 · lost 1 · tied 1 | Call Centers: won; Developers: won; Dictation: tie; Medical: lost; Meetings: won; Self-Hosted: won; Voice Agents: won |
| [Mistral Voxtral Transcribe vs Deepgram](https://www.versusref.com/stt/deepgram-vs-voxtral/) | won 2 · lost 4 · tied 1 | Call Centers: lost; Developers: lost; Dictation: tie; Medical: lost; Meetings: lost; Self-Hosted: won; Voice Agents: won |
| [Mistral Voxtral Transcribe vs ElevenLabs Scribe](https://www.versusref.com/stt/elevenlabs-scribe-vs-voxtral/) | won 3 · lost 2 · tied 1 | Developers: won; Dictation: tie; Medical: won; Meetings: lost; Self-Hosted: won; Voice Agents: lost |

Source: https://www.versusref.com/stt/tools/voxtral/
