Alibaba Unveils Qwen-Audio-3.0-ASR-Flash: A Cutting-Edge Speech Recognition Model Enhancing AI's Grasp of Professional Jargon
14 hour ago / Read about 0 minute
Author:小编   

Alibaba has introduced its advanced speech recognition model, Qwen-Audio-3.0-ASR-Flash. This model has undergone meticulous optimization to ensure contextual coherence, precise recognition of industry-specific terms, and adaptability to emerging terminology. It boasts functionalities for refining speech and generating structured text outputs. By harnessing high-quality glossaries from diverse industries, the model has significantly boosted the recognition accuracy of professional jargon, attaining a remarkable listening accuracy rate of up to 95.36% in medical contexts. The Qwen-Audio-ASR-Flash series has been rigorously tested across various scenarios, including the transcription of meeting minutes, and once claimed the top spot on a global AI evaluation platform with a mere 1.7% typo rate. Presently, Qwen-Audio-3.0-ASR can be accessed and utilized via Alibaba Cloud's BaiLian platform, offering three distinct versions: speech recognition, offline file transcription, and real-time speech recognition, catering to a wide range of user needs.