阿里Qwen-Audio-3.0语音模型上线:识别、合成、实时交互全球第一

凤凰网科技
Yesterday

凤凰网科技讯 8月17日,阿里云今日宣布,Qwen-Audio-3.0系列语音模型正式上线千问AI平台和阿里云百炼平台。开发者可通过API服务调用模型能力,也可订阅Token Plan使用。

该系列包含三款自研语音模型:语音识别大模型Qwen-Audio-3.0-ASR、语音合成大模型Qwen-Audio-3.0-TTS,以及实时语音交互对话模型Qwen-Audio-3.0-Realtime。在全球权威AI评测平台Artificial Analysis今年7月的语音排行榜上,该系列在语音识别、实时交互和语音合成三个赛道均拿下全球第一

其中,ASR模型支持复杂语境和专业领域识别;TTS模型支持多语种、多方言,可控制情绪、语气与节奏;Realtime模型支持端到端语音理解与对话,可边听边说、随时打断追问。目前该系列模型已在千问App、千问办公、Qoder等产品中应用。

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10