SpeechBrain
SpeechBrainAn open-source PyTorch toolkit for training and running speech models, with more than 200 training recipes across about 20 tasks: recognition, synthesis, speaker recognition, diarization, voice activity detection, enhancement, separation, translation and more, and over 100 pretrained models on the Hub.
Openness
5 high confidence- license
- Apache-2.0(OSI)
- source
- public
- core features withheld
- no — an academic, community-run project with no company behind it and no paid tier
SpeechBrain is Apache-2.0 and the published repository is the whole toolkit. It is an academically driven community project rather than a company, and nothing is sold alongside it.
- https://raw.githubusercontent.com/speechbrain/speechbrain/develop/LICENSE recorded 2026-09-27
LICENSE body: "Apache License" / "Version 2.0, January 2004".
- https://raw.githubusercontent.com/speechbrain/speechbrain/develop/README.md recorded 2026-09-27
README: "You are free to redistribute SpeechBrain for both free and commercial purposes, with the condition of retaining license headers."; "SpeechBrain is an academically driven project and relies on the passion and enthusiasm of its contributors."
- https://speechbrain.github.io/ recorded 2026-09-27
Homepage: "SpeechBrain isn't a company or an association. It is an open-source toolkit and a community created by Dr. Mirco Ravanelli and co-created by Dr. Titouan Parcollet."
Adoption
3 high confidencePyPI downloads of the speechbrain package. Its pretrained speaker-embedding model is also pulled directly from the Hub by other tools, which this figure does not count.
- https://pypistats.org/api/packages/speechbrain/recent recorded 2026-09-27
"last_month":979442 for speechbrain.
Capability
4 medium confidenceThe broadest open toolkit for training speech models across tasks, a step below NeMo, whose own recognizers rank near the top of the public boards and which streams.
- https://raw.githubusercontent.com/speechbrain/speechbrain/develop/README.md recorded 2026-09-27
README: "We share over 200 competitive training [recipes](recipes) on more than 40 datasets supporting 20 speech and text processing tasks"; "We are focusing on real-time, streamable, and small-footprint Conversational AI."
Verified 2026-09-27