AI Potluck
Back to Gap Map Model components / Speech & audio

SpeechBrain

SpeechBrain
open source / Overall score: 3.6

An open-source PyTorch toolkit for training and running speech models, with more than 200 training recipes across about 20 tasks: recognition, synthesis, speaker recognition, diarization, voice activity detection, enhancement, separation, translation and more, and over 100 pretrained models on the Hub.

Openness

5 high confidence
5.0
license
Apache-2.0(OSI)
source
public
core features withheld
no — an academic, community-run project with no company behind it and no paid tier

SpeechBrain is Apache-2.0 and the published repository is the whole toolkit. It is an academically driven community project rather than a company, and nothing is sold alongside it.

Adoption

3 high confidence
3.0

PyPI downloads of the speechbrain package. Its pretrained speaker-embedding model is also pulled directly from the Hub by other tools, which this figure does not count.

Capability

4 medium confidence
4.0

The broadest open toolkit for training speech models across tasks, a step below NeMo, whose own recognizers rank near the top of the public boards and which streams.

Verified 2026-09-27