← Developers
Open on GitHub 


#14046 78
Yuekai Zhang
@yuekaizhang · China
1.43k
Weighted score
172
Contributions
18
Repos
Top repos
- espnet/espnet
- k2-fsa/sherpa
- k2-fsa/icefall
- wenet-e2e/wenet
- QwenAudio/CosyVoice
- vllm-project/vllm-omni
- lhotse-speech/lhotse
- NVIDIA-NeMo/RL
- modelscope/FunASR
- NVIDIA-NeMo/Automodel
Ranked AI repos3
126689
QwenAudio/CosyVoice
+7
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
23.9k· Python· Models
4522246
FireRedTeam/FireRedASR
+1
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
2.00k· Python· Models
33127863
k2-fsa/sherpa
+2
Speech-to-text server framework with next-gen Kaldi
1.00k· C++· Infrastructure