AI Potluck
Back to Gap Map Model components / Language-specific datasets

INCLUDE

Cohere
open / Overall score: 3.7

INCLUDE is a multilingual knowledge and reasoning benchmark of four-option multiple-choice questions taken from local academic, professional and licensing exams in 44 languages. Because the questions come from each region's own exams rather than translated English tests, many probe regional knowledge. A smaller lite subset covers the same languages. It was built by EPFL and Cohere Labs researchers.

Openness

5 high confidence
5.0
license
apache-2.0
access
public
dataset_card
present

Both repositories download from Hugging Face without a gate under Apache-2.0, and the full question set including answers is published.

Adoption

3 high confidence
3.0

Hugging Face downloads over the trailing month for the base and lite repositories combined. Runs through evaluation harnesses that cache the data are counted only when they download it.

Capability

4 medium confidence
4.0

INCLUDE draws its questions from each region's own exams rather than translated English tests, is documented in its paper, and ships as a task in EleutherAI's evaluation harness. Its 44 languages are fewer than the 122 variants of Belebele.

Verified 2026-09-24