AI Potluck
Back to Gap Map Model components / Speech & audio

MeloTTS

MyShell
open weights / Overall score: 1.8

A multilingual text-to-speech library from MyShell and MIT, with voices for several English accents, Spanish, French, Chinese mixed with English, Japanese and Korean, fast enough for real-time synthesis on a CPU.

The repository has not been pushed to since December 2024 and the newest checkpoint dates from April 2024. A PyPI package named melotts exists but names no repository, so it is not declared.

Openness

3 high confidence
3.0
weights
open(eight per-language .pth checkpoints on the Hub, ungated)
data
closed(no statement of what the released checkpoints were trained on)
code
open(training scripts in the repository (train.sh, preprocess_text.py))
license
MIT(code and every checkpoint)

MeloTTS is MIT-licensed, code and checkpoints alike, and the repository includes the scripts to train a voice on your own data. What the released voices were trained on is not stated, so they cannot be rebuilt.

Adoption

3 high confidence
3.0

Hugging Face downloads summed over the per-language checkpoints, about half of them the English voice.

Capability

1 low confidence
1.0

A light, practical synthesizer for fixed voices. It clones no voices and publishes no quality results, and listeners have not rated it on the public arena, so it sits well below Kokoro.

Verified 2026-09-27