AI Potluck
Back to Gap Map Model components / Language-specific datasets

DarijaMMLU

Mohamed bin Zayed University of Artificial Intelligence
open / Overall score: 2.0

DarijaMMLU tests language models on 22,027 multiple-choice questions in Moroccan Darija across 44 subjects. The questions are selected subsets of MMLU and ArabicMMLU, machine-translated into Darija with Claude 3.5 Sonnet and then manually reviewed and adapted, with technical and culturally inappropriate topics left out. MBZUAI-Paris built it as part of the Atlas-Chat evaluation suite.

Openness

5 medium confidence
5.0
license
mit(card metadata
access
public(ungated on the Hub)
dataset_card
present(card describes sources, translation model and review)

The questions download freely and the card states where each one came from and how it was translated. Whether the permissive grant can cover the ArabicMMLU-derived questions is unclear, since the source benchmark forbids commercial use.

Adoption

2 high confidence
2.0

Hugging Face downloads of the Hub repository. Evaluation harnesses fetch a benchmark on each run, so the figure counts runs, not the groups using it.

Capability

2 medium confidence
2.0

DarijaMMLU gives Moroccan Darija a broad knowledge benchmark and is the main test for the Atlas-Chat models. Its questions are machine translations of MMLU and ArabicMMLU with manual review, not material written in Darija, whereas KazMMLU is built from native educational material.

Verified 2026-09-24