Not established from the available sources.
GOOGLE (DEEPMIND) · MODEL RELEASE TRACKER
Gemini 3.8 Flash-Lite TTS
Gemini 3.8 Flash-Lite TTS is a text-to-speech (audio generation) model in the Gemini 3.8 family, introduced September 23, 2026. It is optimized for high-volume, cost-efficient speech generation (dubbing, audio content, voice agents), supports features like stage directions, two-voice dialogue, and nonverbal sounds, and is available through Google AI Studio, the Gemini API and related Google platforms. Generated clips include an inaudible SynthID watermark.CURRENT SNAPSHOT2/5 DIMENSIONS WITH DATA
The dimensions that change the decision.
Not established from the available sources.
Not established from the available sources.
RELEASE TIMELINE
Published, source-backed release events only.
Google DeepMind unveils Gemini 4 Argon, a frontier model with 1M-token context
Google DeepMind announced Gemini 4 Argon, a new frontier multimodal model optimized for long-horizon reasoning and complex workflows. Argon is rolling out initially to trusted cyber defenders via the Fairwind Program, expands context length to 1 million tokens, reports leading benchmark performance across coding, finance, legal, and video understanding, and will be made more widely available after phased safety testing and engagement with U.S. government pre-release processes; Google also published introductory pricing.
Google releases Gemini 3.8 Flash and Flash‑Lite text‑to‑speech models
Google introduced two new Gemini 3.8 text‑to‑speech models—Flash TTS for highly expressive, character-driven voice design and Flash‑Lite TTS for cost‑efficient, high‑volume use—available across Google AI Studio, Gemini API/Enterprise, Gemini Notebook, and Google Vids. The models claim large voice coverage (2,000+ production voices, 100+ languages/dialects), generative voice creation from prompts, 30‑second voice replication with consent verification, SynthID watermarking and C2PA credentials, and features for long‑form and multi‑speaker scene staging.
VERIFIABLE FACTS
Every value stays attached to a source and date.
Announced September 23, 2026.
Text-to-speech (audio generation) model.
Built for high-volume, cost-efficient scale; optimized for high-volume dubbing, audio content creation, and expressive voice agents with fine-grained control over tone, pacing, and expressive nuance.
Available through Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids (per Google posts).
Rolling out through Gemini API and Google AI Studio; Gemini Enterprise API access will follow (per The Decoder report).
Supports more than 100 languages and dialects.
Supports stage directions per line, two-voice dialogue, and nonverbal sounds such as laughter and sighs.
Generated audio clips include an inaudible SynthID watermark to help detect AI-generated speech (per The Decoder quoting Google).
WHAT CHANGED