Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

MODEL · ENTITY #5214

Qwen Cloud

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

MODELS · 1 SOURCE · The Decoder

Alibaba releases Qwen-Audio-3.1 suite and cuts AI audio prices up to 95%

Alibaba's Qwen team released Qwen-Audio-3.1, a set of five models for automatic speech recognition (ASR), text‑to‑speech (TTS) and real‑time interaction. New capabilities include improved multilingual and dialect ASR, multi‑speaker timestamps and emotion/ambient/machine‑noise detection in ASR-Next, cross‑language and controllable TTS and a diffusion‑based TTS-Next that generates voice, effects and background audio in one pass; the real‑time model supports simultaneous speaking/listening with instant interruption and mood‑aware responses. Alibaba also sharply cut prices on its audio services (TTS ≈70% lower, Realtime ≈85% lower, ASR up to 95%) and made details available via its blog and Qwen Cloud.

7.0

MODELS · 1 SOURCE · The Decoder

Qwen launches Qwen3.8-Omni-Flash multimodal agent model and undercuts Gemini Flash pricing

Qwen announced Qwen3.8-Omni-Flash, its first multimodal agent-focused model with a one‑million‑token context window that processes audio and video together, performs tasks like vlog editing and movie summarization, and supports real‑time interaction via Qwen‑Live Harness. API pricing is $0.15 per million input tokens and $0.47 per million output tokens (Qwen estimates audio input under $0.01/hour and 720p video at 1 fps about $0.20/hour), and the model is available through Qwen Studio, Qwen Cloud and the Qwen API; Qwen also published open-source Qwen-MM-Plugins for agent integrations.

8.0