modelsSEP 10 04:00 UTC
Qwen-Audio-3.0-ASR technical report details LLM-integrated speech recognition
A technical report published on arXiv introduces Qwen-Audio-3.0-ASR, an automatic speech recognition system that combines scaled training data, larger model architecture, and integration with large language models. The paper outlines the system's design choices and evaluates its performance, situating it within recent progress in ASR research.