← Open Source
k2-fsa

sherpa

Speech-to-text server framework with next-gen Kaldi

InfrastructureDeploy & ServeC++
Open on GitHub
Momentum
+0stars in 24 hours0.0%
999
Stars
151
Forks
+2
This week
58
Contributors
Created 2022-05-20 · Updated 2026-10-06 · #11397 today
Top developers
README

sherpa

sherpa is an open-source speech-text-text inference framework using PyTorch, focusing exclusively on end-to-end (E2E) models, namely transducer- and CTC-based models. It provides both C++ and Python APIs.

This project focuses on deployment, i.e., using pre-trained models to transcribe speech. If you are interested in how to train or fine-tune your own models, please refer to icefall.

We also have other similar projects that don't depend on PyTorch:

sherpa-onnx and sherpa-ncnn also support iOS, Android and embedded systems.

Installation and Usage

Please refer to the documentation at

Try it in your browser

Try sherpa from within your browser without installing anything: