Tech Meridian← ALL MODELS
PROMY MERIDIAN

MISTRAL AI · MODEL RELEASE TRACKER

Voxtral TTS

Voxtral TTS is a 4B-parameter text-to-speech model from Mistral AI (announced 2026-03-23) that targets multilingual, emotionally expressive voice generation in 9 languages with support for diverse dialects. The model emphasizes contextual understanding, speaker modeling, zero-shot cross-lingual voice adaptation, low latency, and easy voice adaptation. It is offered via API and Mistral Studio with pricing starting at $0.016 per 1k characters.

CURRENT SNAPSHOT4/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING
PricingPricing starts at $0.016 per 1k characters.
CONTEXT WINDOW

Not established from the available sources.

MODALITIES
Modality and language supportText-to-speech model providing multilingual voice generation in 9 languages with support for diverse dialects.
BENCHMARKS
Performance claimDescribed by Mistral as delivering "state-of-the-art" performance in multilingual voice generation; mentions automated metrics such as word-error-rate.
AVAILABILITY
AvailabilityAvailable via API and Mistral Studio (Mistral Voices).

VERIFIABLE FACTS

Every value stays attached to a source and date.

RELEASE · Release / relation to developerDEVELOPER CLAIM

Announced by Mistral AI on 2026-03-23; described as Mistral's first text-to-speech model.

MODALITIES · Modality and language supportDEVELOPER CLAIM

Text-to-speech model providing multilingual voice generation in 9 languages with support for diverse dialects.

CAPABILITIES · Expressiveness, contextual understanding, and voice adaptationDEVELOPER CLAIM

Claims realistic, emotionally expressive speech; excels at contextual understanding and speaker modeling; supports easy voice adaptation and zero-shot cross-lingual voice adaptation.

CAPABILITIES · Latency and intended use casesDEVELOPER CLAIM

Described as lightweight with low latency and suitable for enterprise voice workflows and scalable AI agents.

BENCHMARKS · Performance claimDEVELOPER CLAIM

Described by Mistral as delivering "state-of-the-art" performance in multilingual voice generation; mentions automated metrics such as word-error-rate.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport created8 facts