Tech Meridian← ALL MODELS
PROMY MERIDIAN

MISTRAL AI · MODEL RELEASE TRACKER

Voxtral Realtime

Voxtral Realtime is a streaming speech-to-text model in the Voxtral Transcribe 2 family, purpose-built for live transcription with configurable latency (down to sub-200ms). It is released with open weights under the Apache 2.0 license and is optimized to run on edge devices (4B parameters).

CURRENT SNAPSHOT3/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING

Not established from the available sources.

CONTEXT WINDOW

Not established from the available sources.

MODALITIES
ModalitySpeech-to-text (streaming / live transcription).
BENCHMARKS
Benchmarks and reported performanceEvaluated on the FLEURS transcription benchmark and multiple diarization benchmarks (Switchboard, CallHome, AMI-IHM, AMI-SDM, SBCSAE, TalkBank). Reported: at 2.4s delay Realtime matches Voxtral Mini Transcribe V2; at 480ms delay it stays within ~1–2% word error rate.
AVAILABILITY
Weights and license / availabilityVoxtral Realtime ships with open weights under the Apache 2.0 license and the weights are released on the Hugging Face Hub; stated as deployable on edge.

VERIFIABLE FACTS

Every value stays attached to a source and date.

BENCHMARKS · Benchmarks and reported performanceDEVELOPER CLAIM

Evaluated on the FLEURS transcription benchmark and multiple diarization benchmarks (Switchboard, CallHome, AMI-IHM, AMI-SDM, SBCSAE, TalkBank). Reported: at 2.4s delay Realtime matches Voxtral Mini Transcribe V2; at 480ms delay it stays within ~1–2% word error rate.

CAPABILITIES · Multilingual supportDEVELOPER CLAIM

Natively multilingual with strong transcription performance in 13 languages (including English, Chinese, Hindi, Spanish, Arabic, French, Portuguese, Russian, German, Japanese, Korean, Italian, Dutch).

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport created8 facts