Tech Meridian← ALL MODELS
PROMY MERIDIAN

ALIBABA (QWEN) · MODEL RELEASE TRACKER

Qwen-Audio-3.1

Qwen-Audio-3.1 is a released lineup of five audio models from Alibaba's Qwen team for automatic speech recognition (ASR), text-to-speech (TTS), and real-time interaction, introducing new features and substantial price cuts.

CURRENT SNAPSHOT0/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING

Not established from the available sources.

CONTEXT WINDOW

Not established from the available sources.

MODALITIES

Not established from the available sources.

BENCHMARKS

Not established from the available sources.

AVAILABILITY

Not established from the available sources.

RELEASE TIMELINE

Published, source-backed release events only.

1 SOURCES · IMPORTANCE 7.0

Alibaba releases Qwen-Audio-3.1 suite and cuts AI audio prices up to 95%

Alibaba's Qwen team released Qwen-Audio-3.1, a set of five models for automatic speech recognition (ASR), text‑to‑speech (TTS) and real‑time interaction. New capabilities include improved multilingual and dialect ASR, multi‑speaker timestamps and emotion/ambient/machine‑noise detection in ASR-Next, cross‑language and controllable TTS and a diffusion‑based TTS-Next that generates voice, effects and background audio in one pass; the real‑time model supports simultaneous speaking/listening with instant interruption and mood‑aware responses. Alibaba also sharply cut prices on its audio services (TTS ≈70% lower, Realtime ≈85% lower, ASR up to 95%) and made details available via its blog and Qwen Cloud.

→

VERIFIABLE FACTS

Every value stays attached to a source and date.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport created8 facts