AI Potluck
Back to Gap Map Model components / Speech & audio

sherpa-onnx

k2-fsa (Next-gen Kaldi)
open source / Overall score: 4.0(strong)

The Next-gen Kaldi project's ONNX Runtime engine for offline speech: streaming and non-streaming recognition, synthesis, speaker diarization and identification, voice activity detection and more, on Linux, Windows, macOS, Android, iOS, HarmonyOS, embedded boards and WebAssembly, with bindings for a dozen programming languages.

Openness

5 medium confidence
5.0
license
Apache-2.0(OSI)
source
public
core features withheld
no

sherpa-onnx is Apache-2.0 and the published repository is the whole engine. No maintainer statement about commercial offerings was found, so that reading comes from the repository alone.

Adoption

4 high confidence
4.0

PyPI downloads of sherpa-onnx. Its mobile and native builds ship outside PyPI and are not counted.

Capability

4 medium confidence
4.0

The broadest on-device speech engine here, serving most speech tasks, but an inference engine rather than a training toolkit.

Verified 2026-09-27