← Developers
Open on GitHub 


#13940 88
Yuekai Zhang
@yuekaizhang · China
1.42k
Weighted score
171
Contributions
17
Repos
Top repos
- espnet/espnet
- k2-fsa/sherpa
- k2-fsa/icefall
- wenet-e2e/wenet
- QwenAudio/CosyVoice
- vllm-project/vllm-omni
- lhotse-speech/lhotse
- NVIDIA-NeMo/RL
- modelscope/FunASR
- NVIDIA-NeMo/Automodel
Ranked AI repos3
1153304
QwenAudio/CosyVoice
+7
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
23.8k· Python· Models
7573746
FireRedTeam/FireRedASR
+0
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
2.00k· Python· Models
102471374
k2-fsa/sherpa
+0
Speech-to-text server framework with next-gen Kaldi
998· C++· Infrastructure