Meta甩出语音转写“神器”:准确识别中英文夹杂,支持超20人长聊1小时

智东西
Sep 02

语音转写正变成AI系统的实时感知层。

编译 | ZeR0

编辑 | 漠影

智东西9月2日报道,今日,Meta推出由Meta超级智能实验室开发的首个实时音频感知模型Muse Voice Transcribe

该模型提供实时流式自动语音识别(ASR)、支持超过20位说话人的语音分割以及端点控制功能,支持多语言,能处理超过1小时的长音频,可实现无缝语码切换,并通过语言、关键词和上下文偏好来提高识别准确率。

从示例来看,无论是讲中文、中式英语口音,还是中英文混合表达,这款模型的转写速度和效果都相当不错。

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10