Qwen-Audio-3.1 lower Qwen-Audio full voice model price

It is known that the Qwen-Audio-3.1 series of large voice models were officially published. This upgrade not only fully evolved the three core models of speech recognition (ASR), speech synthesis (TTS) and real-time voice interaction (Realtime), but also introduced a new audio creation model Qwen-Audio-31-TS-Next and audio interpretation model Qwen-Audio-31-ASR-Next at a heavier pound. Five completely new voice models were launched at the same time, resulting in a complete set of audio-capacity stacks covering “understanding-generation-interactive-creation”. To further reduce user usage costs, the prices of the Qwen-Audio full-line voice model were revised downwards, with TTS down by about 70%, Realtime down by about 85% and ASR down by 95%。

Search