← 开源
k2-fsa

sherpa

Speech-to-text server framework with next-gen Kaldi

InfrastructureDeploy & ServeC++
在 GitHub 打开
增长势头
+124 小时新增 Star+0.1%
999
Star
151
Fork
+2
本周
58
贡献者
创建于 2022-05-20 · 更新于 2026-10-06 · 今日第 5817 名
主要开发者
README

sherpa

sherpa is an open-source speech-text-text inference framework using PyTorch, focusing exclusively on end-to-end (E2E) models, namely transducer- and CTC-based models. It provides both C++ and Python APIs.

This project focuses on deployment, i.e., using pre-trained models to transcribe speech. If you are interested in how to train or fine-tune your own models, please refer to icefall.

We also have other similar projects that don't depend on PyTorch:

sherpa-onnx and sherpa-ncnn also support iOS, Android and embedded systems.

Installation and Usage

Please refer to the documentation at

Try it in your browser

Try sherpa from within your browser without installing anything: