Khayyam Challenge
Raia CenterThe Khayyam Challenge, also called PersianMMLU, tests language models on 20,192 four-choice questions in Persian from 38 tasks drawn from Persian examinations, from lower primary to upper secondary school. The questions are original rather than translated and carry metadata such as human response rates, difficulty ratings, and descriptive answers. The RAIA Center publishes it.
The Hub README is empty; the description follows the paper and the gating terms.
Openness
2 high confidence- license
- cc-by-nd-4.0(card metadata
- access
- manual(manual approval after a form sharing contact information and accepting the terms)
- dataset_card
- no(README.md exists but is empty)
Access needs a form and manual approval, and the terms allow only non-commercial academic research and forbid derivative works, including new benchmarks built from the questions. The repository has no card, so its contents are described only in the paper. The no-derivatives term alone would keep it from being open data, since a model trained on the questions is arguably a derivative.
- https://huggingface.co/api/datasets/raia-center/khayyam-challenge recorded 2026-09-24
API JSON: "gated": "manual"; tag license:cc-by-nd-4.0.
- https://huggingface.co/datasets/raia-center/khayyam-challenge recorded 2026-09-24
Gating terms: "distributed under a Creative Commons No Derivatives (CC ND) license, prohibiting the creation of derivative works ... designated exclusively for non-commercial, academic research"; "README.md exists but content is empty."
Adoption
1 high confidenceHugging Face downloads of the gated repository. Each approved user fetches it directly, so the figure is closer to a count of users than for ungated sets, but copies inside leaderboards are not counted.
- https://huggingface.co/api/datasets/raia-center/khayyam-challenge recorded 2026-09-24
43 downloads in the trailing 30 days for raia-center/khayyam-challenge
Capability
4 medium confidenceThe Khayyam Challenge is a documented Persian knowledge benchmark of original exam questions with human response rates, and the ParsBench leaderboard uses it. It is a step below ArabicMMLU because it has no card and carries restrictive terms, which make it harder to adopt.
- https://arxiv.org/abs/2404.06644 recorded 2026-09-24
Abstract: "Khayyam Challenge (also known as PersianMMLU), a meticulously curated collection comprising 20,192 four-choice questions sourced from 38 diverse tasks extracted from Persian examinations".
- https://huggingface.co/datasets/raia-center/khayyam-challenge recorded 2026-09-24
Hub page: "Space using raia-center/khayyam-challenge 1: ParsBench/leaderboard".
Verified 2026-09-24