NEWS · MODELS · #454
Grok Voice Think Fast 2.0 released: vendor announces next‑gen speech-to-speech model with higher accuracy and reasoning
The vendor announced Grok Voice Think Fast 2.0, a next‑generation speech-to-speech model that it says improves transcription accuracy, conversational behavior, and on-the-fly reasoning versus Grok Voice Think Fast 1.0 and compared competitors (Deepgram Nova 3, ElevenLabs Scribe v2). The announcement reports 1.4×–2.0× gains on thousands of short phrases, ~10× advantage in noisy settings, A/B improvements in Starlink support/sales tests, rollout to grok-voice-latest on August 5, 2026, and pricing of $0.08 per minute of audio.
KEY POINTS
- The vendor announced Grok Voice Think Fast 2.0, a next‑generation speech-to-speech model that it says improves transcription accuracy, conversational behavior, and on-the-fly reasoning versus Grok Voice Think Fast 1.0 and compared competitors (Deepgram Nova 3, ElevenLabs Scribe v2).
- The announcement reports 1.4×–2.0× gains on thousands of short phrases, ~10× advantage in noisy settings, A/B improvements in Starlink support/sales tests, rollout to grok-voice-latest on August 5, 2026, and pricing of $0.08 per minute of audio.
- A vendor claim of a speech-to-speech model that reasons while speaking and outperforms dedicated transcription systems — with production testing and a public rollout date and price — could shift adoption and competitive dynamics in voice AI and contact-center automation.
WHY IT MATTERS
A vendor claim of a speech-to-speech model that reasons while speaking and outperforms dedicated transcription systems — with production testing and a public rollout date and price — could shift adoption and competitive dynamics in voice AI and contact-center automation.