Press Enter to search · ESC to close

AI × Crypto

Alibaba’s Qwen Releases Qwen-Audio-3.1, Cuts Speech Model Prices by Up to 95%

Alibaba’s Qwen released the Qwen-Audio-3.1 series of speech models, upgrading its ASR, TTS and Realtime voice interaction models and adding the new Qwen-Audio-3.1-TTS-Next audio creation model and Qwen-Audio-3.1-ASR-Next audio understanding model. Five new speech models were launched together, covering understanding, generation, interaction and creation. Prices across the Qwen-Audio line were cut, with TTS down about 70%, Realtime about 85% and ASR 95%.

Original source

AI take

The headline number is the price cut, but the more consequential move is the breadth of the release: understanding, generation, interaction and creation models shipped together, which turns speech from a single API call into a stack. That matters most for developers building voice agents, where ASR and TTS costs have been a real constraint on always-on interaction. Whether rivals match these prices, and whether cheap speech plus capable models pulls more builders into voice-first products, is the open question.

Generated by AI for reference only.

Share

Related News

TREE NEWS share card
Long-press image above → Save to Photos / Share
Pitch us Feedback