Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

COMPANY · ENTITY #8173

WhisperX

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

CODING · 1 SOURCE · AWS Machine Learning

Deploy streamed Qwen3-TTS on SageMaker using vLLM-Omni v1.5 for real-time voice

AWS published a Part 1 tutorial showing how to deploy the Qwen3-TTS text-to-speech model on Amazon SageMaker AI using the vLLM-Omni v1.5 Deep Learning Container (DLC). The guide demonstrates sending text and receiving audio chunks over a single persistent bidirectional connection (SageMaker bidirectional streaming) via the vLLM-Omni native WebSocket route (v1/audio/speech/stream) and includes a Gradio example; the series will also cover other specialized DLCs like WhisperX and llama.cpp.

6.0

MODELS · 1 SOURCE · AWS Machine Learning

AWS packages WhisperX DLC for SageMaker to deliver speaker‑labeled, per‑word timestamps

AWS provides a WhisperX Deep Learning Container (DLC) for Amazon SageMaker that bundles OpenAI Whisper with wav2vec2 forced alignment and speaker diarization to produce per-word timestamps and speaker labels. The GPU-ready container conforms to SageMaker’s serving contract, supports real-time and asynchronous endpoints, and outputs json, verbose_json, srt, and vtt formats (example image tag: whisperx:3.8.6-cu128-amzn2023-sagemaker).

6.0