AI Potluck
Back to Gap Map Model components / Language-specific datasets

LaoBench

Beijing Academy of Artificial Intelligence (BAAI)
gated / Overall score: 2.4

LaoBench is a benchmark for testing language models in Lao, with more than 17,000 expert-curated items across three areas: culturally grounded knowledge, K-12 curriculum questions, and translation among Lao, Chinese and English. Experts wrote the items, and an agent-assisted pipeline verified them. It comes in an open subset and a held-out subset scored through a controlled service. BAAI publishes it.

The Hugging Face card carries only metadata; the description and scale come from the LaoBench paper.

Openness

2 medium confidence
2.0
license
apache-2.0(card metadata
access
public(open subset ungated
dataset_card
partial(README has YAML metadata only, no prose)
answers
held-out(paper: held-out portion enables black-box evaluation via a controlled service

The open subset downloads without a gate under Apache-2.0, while part of the benchmark is kept private and scored by BAAI's service. The repository card has no description, so how the data was made is documented only in the paper.

Adoption

1 high confidence
1.0

Adoption is Hugging Face downloads of the open subset. Evaluations on the held-out portion go through BAAI's service and are not counted.

Capability

3 medium confidence
3.0

LaoBench is documented in a paper and offers expert-written Lao items in knowledge, schooling and translation. No outside leaderboard or model report using it was found, less than half of it is openly downloadable, and SEA-HELM tests Lao among eleven Southeast Asian languages on its own leaderboard.

Verified 2026-09-24