AI Potluck
Back to Gap Map Model components / Language-specific datasets

ZA African Next Voices

Data Science for Social Impact, University of Pretoria
gated / Overall score: 2.7

ZA African Next Voices, also called Swivuriso, is a speech recognition dataset of about 3,000 hours of scripted and unscripted speech in seven South African languages: isiZulu, isiXhosa, Sesotho, Setswana, Xitsonga, isiNdebele and Tshivenda. Topics span agriculture, health, finance and general subjects, with speaker-disjoint splits. The South Africa Next Voices team publishes it under the dsfsi-anv Hugging Face organization, alongside a compressed copy.

Openness

3 high confidence
3.0
license
cc-by-4.0(card adds a use restriction prohibiting TTS, voice cloning and voice synthesis)
access
auto(automatic Hugging Face gate)
dataset_card
present(languages, hours, fields, splits and use restriction described)

The audio carries an attribution-only license behind an automatic click-through, but the card forbids any use for speech synthesis or voice cloning, which narrows what that license alone would allow. A test portion is held back for future shared tasks.

Adoption

2 high confidence
2.0

Hugging Face downloads summed over the full repository and its Opus-compressed copy of the same audio. Downloads count file fetches behind an automatic gate, not users.

Capability

3 medium confidence
3.0

ZA African Next Voices gives seven South African languages scripted and unscripted speech transcribed by people, five hundred hours each for five of them, with a paper on its design, and its publishers' Whisper models and a community isiZulu recognizer are trained on it. Its roughly three thousand hours are more than any other open set here offers these languages, yet a small share of the hundreds of thousands behind the largest English speech corpora.

Verified 2026-09-24