AI Potluck
Back to Gap Map Model components / Language-specific datasets

MILU

AI4Bharat
gated / Overall score: 2.7

MILU (Multi-task Indic Language Understanding Benchmark) is an evaluation set of about 80,000 multiple-choice questions in 10 Indic languages and English, spanning 8 domains and 41 subjects. Its questions come largely from regional and state examinations, about a quarter of them translated, and cover India-specific topics such as local history, arts and law. AI4Bharat publishes it with an lm-evaluation-harness setup for running it.

Openness

3 high confidence
3.0
license
cc-by-4.0(card metadata and README "License" section
access
auto(Hugging Face gated: auto)
dataset_card
present(card gives per-language statistics, subjects and splits)

The benchmark carries CC BY 4.0, and the Hugging Face copy needs an access request that is approved automatically. The MIT license file in the GitHub repository belongs to the EleutherAI harness code, not the questions.

Adoption

2 high confidence
2.0

Hugging Face downloads of the MILU repository. Evaluation runs through the harness also fetch from here, but a download does not show which published results used it.

Capability

3 medium confidence
3.0

MILU is a documented knowledge benchmark built from Indian exam questions in 11 languages, and its paper evaluates more than 40 models on it. No leaderboard or model report outside that paper was found, whereas IndicXTREME was released with the IndicBERT v2 models it tests.

Verified 2026-09-24