AI Potluck
Back to Gap Map Model components / Language-specific datasets

Belebele

Meta
open / Overall score: 4.4(strong)

Belebele is a multiple-choice reading comprehension benchmark in 122 language variants, with 900 questions per variant written about short FLORES-200 passages. Every question has four answers, one correct, and the set is fully parallel so scores compare directly across languages. Annotators wrote the questions in English under quality checks before translation. Meta FAIR released it.

Openness

5 high confidence
5.0
license
cc-by-sa-4.0(Covers the benchmark
access
public(Ungated on Hugging Face and as a zip from Meta's file server.)
dataset_card
present

The benchmark carries a share-alike license and downloads without a gate. The optional English training set assembled from other datasets is mostly non-commercial, but that part is not the benchmark.

Adoption

3 high confidence
3.0

Hugging Face downloads over the trailing month for the Belebele repository. Downloads of the zip from Meta's file server are not counted.

Capability

5 medium confidence
5.0

Belebele asks the same 900 questions in 122 language variants, so reading comprehension scores compare directly across languages, and no other open reading test here reaches that many. It is documented in its paper and ships as a task in EleutherAI's evaluation harness.

Verified 2026-09-24