IndoNLU
IndoNLPIndoNLU is a benchmark for Indonesian natural language understanding with twelve tasks, from emotion and sentiment classification to part-of-speech tagging, named-entity recognition, keyphrase extraction, textual entailment and extractive QA. The task data comes from social media, reviews, news and Wikipedia, some of it labeled by Indonesian linguists. It shipped with the Indo4B pretraining corpus and the IndoBERT models, from a collaboration including Institut Teknologi Bandung, Universitas Indonesia, Gojek and Prosa.AI.
Openness
2 medium confidence- license
- mit(HF card: 'The licensing status of the IndoNLU benchmark datasets is under MIT License'
- access
- public(ungated)
- dataset_card
- present(per-task descriptions
- answers
- held-out(GitHub test-set labels are masked
The task data downloads freely under a permissive license, but the official test labels are masked and scored through a CodaLab competition. The Hugging Face card leaves its collection and annotation sections unfilled, so provenance comes from the paper.
- https://huggingface.co/api/datasets/indonlp/indonlu recorded 2026-09-24
gated: false; license:mit
- https://huggingface.co/datasets/indonlp/indonlu/raw/main/README.md recorded 2026-09-24
"The licensing status of the IndoNLU benchmark datasets is under MIT License."; Source Data: "[Needs More Information]"
- https://raw.githubusercontent.com/IndoNLP/indonlu/master/LICENSE recorded 2026-09-24
Apache License, Version 2.0 (repository LICENSE)
- https://raw.githubusercontent.com/IndoNLP/indonlu/master/README.md recorded 2026-09-24
"The labels of the test set are masked (no true labels) in order to preserve the integrity of the evaluation. Please submit your predictions to the submission portal at CodaLab"
Adoption
1 high confidenceAdoption is Hugging Face downloads of the repository. Users of the GitHub copies and of the CodaLab test sets are not counted.
- https://huggingface.co/api/datasets/indonlp/indonlu recorded 2026-09-24
571 downloads in the trailing 30 days for indonlp/indonlu
Capability
4 medium confidenceIndoNLU is documented in a paper, shipped with the IndoBERT models, and many Indonesian classifiers are fine-tuned on its tasks. It covers only Indonesian and only understanding tasks, while SEA-HELM tests Indonesian alongside other Southeast Asian languages and adds generative, cultural and safety tasks.
- https://arxiv.org/abs/2009.05387 recorded 2026-09-24
"we introduce the first-ever vast resource for the training, evaluating, and benchmarking on Indonesian natural language understanding (IndoNLU) tasks. IndoNLU includes twelve tasks"
- https://huggingface.co/datasets/indonlp/indonlu recorded 2026-09-24
Models trained or fine-tuned on indonlp/indonlu: w11wo/indonesian-roberta-base-posp-tagger, w11wo/indonesian-roberta-base-sentiment-classifier
- https://huggingface.co/datasets/indonlp/indonlu/raw/main/README.md recorded 2026-09-24
"SmSA ... The text was crawled and then annotated by several Indonesian linguists"
Verified 2026-09-24