Read the original: FunAudioLLM (Alibaba Tongyi) · new models on Hugging Face·Published AI score36/100
Fun-ASR-Nano-2512 Speech Recognition Model Released by Tongyi Lab on Hugging Face
Original titleFunAudioLLM/Fun-ASR-Nano-2512
AISummary
Tongyi Lab has released Fun-ASR-Nano-2512, an end-to-end speech recognition large model trained on tens of millions of hours of real speech, supporting low-latency real-time transcription across 31 languages.
The model, which has 800M parameters, targets industry use such as education and finance and claims 93% accuracy in far-field, high-noise conditions. It is available on Hugging Face and works with the FunASR toolkit.
Source: FunAudioLLM (Alibaba Tongyi) · new models on Hugging Face · huggingface.co