Vuk'uzenzele isiXhosa Speech Dataset (ViXSD)
Lelapa AIThe Vuk'uzenzele isiXhosa Speech Dataset (ViXSD) is about 10 hours of read isiXhosa speech in which eight native speakers, four men and four women, narrate articles from the Vuk'uzenzele corpus. Each of its 395 recordings comes with a transcript and speaker metadata on age, accent and linguistic background. Lelapa AI released it as the proof of concept for its community-centered Esethu Framework and license.
Openness
3 high confidence- license
- Esethu-License(card
- access
- auto(automatic Hugging Face gate)
- dataset_card
- present(contents, speakers, fields, splits and a datasheet)
It opens after an automatic click-through and is governed by the Esethu License, a community-centric data license designed to share benefits with the speakers' community, rather than a standard open license. African entities may use it commercially for free and other commercial users pay a fee, so commercial use is possible but not on equal terms for everyone.
- https://arxiv.org/abs/2502.15916 recorded 2026-09-24
Abstract: "the Esethu license, a novel community-centric data license"; ViXSD introduced "as a proof of concept".
- https://huggingface.co/api/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD recorded 2026-09-24
"gated": "auto"; cardData license "other".
- https://huggingface.co/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD recorded 2026-09-24
"License: Esethu License"; "395 stereo audio recordings... a total of 10 hours of narrated speech isiXhosa narrated by 8 speakers".
Adoption
1 high confidenceHugging Face downloads of the single repository, behind an automatic gate. Downloads count file fetches, not users.
- https://huggingface.co/api/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD recorded 2026-09-24
27 downloads in the trailing 30 days for lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD
Capability
1 medium confidenceViXSD gives isiXhosa read speech from eight native speakers, documented with a paper, speaker metadata and a datasheet, though nothing beyond its own paper's recognition experiments uses it. At about ten hours it is a minute sample beside English speech corpora of hundreds of thousands of hours, alongside BibleTTS as a documented speech set with no outside use.
- https://huggingface.co/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD recorded 2026-09-24
"It contains a total of 10 hours of narrated speech isiXhosa narrated by 8 speakers (4 male, 4 female) with approximately 39,000 words."
Verified 2026-09-24