Skip to content

Fun-ASR-Nano-2512 Speech Recognition Model Released by Tongyi Lab on Hugging Face

Original titleFunAudioLLM/Fun-ASR-Nano-2512

AISummary

Tongyi Lab has released Fun-ASR-Nano-2512, an end-to-end speech recognition large model trained on tens of millions of hours of real speech, supporting low-latency real-time transcription across 31 languages.

The model, which has 800M parameters, targets industry use such as education and finance and claims 93% accuracy in far-field, high-noise conditions. It is available on Hugging Face and works with the FunASR toolkit.

Read the original huggingface.co

Source: FunAudioLLM (Alibaba Tongyi) · new models on Hugging Face · huggingface.co