datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mlaire-xquad
MLAIRE-XQUAD
XQuAD reformatted for language-aware retrieval evaluation. Each passage appears once per language; relevance is encoded by group_id matching.
This repository is part of the MLAIRE benchmark, submitted anonymously
to the NeurIPS 2026 Evaluations & Datasets Track. Authors and affiliations
are withheld for double-blind review.
Default top-k
Reported metrics in the paper use top-20.
Layout
corpus/test-*.parquet _id, text, title, language, group_id… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-submission-nips/mlaire-xquad.mlaire-belebele
MLAIRE-BELEBELE
Belebele reformatted for language-aware retrieval evaluation. 488 underlying passages, each available in 122 languages (joined globally by the original link field). Relevance is encoded by group_id matching.
This repository is part of the MLAIRE benchmark, submitted anonymously
to the NeurIPS 2026 Evaluations & Datasets Track. Authors and affiliations
are withheld for double-blind review.
Default top-k
Reported metrics in the paper use top-200.… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-submission-nips/mlaire-belebele.mlaire-mlqa
MLAIRE-MLQA
MLQA reformatted for language-aware retrieval evaluation. Passages are deduplicated at the context level via union-find on the original MLQA ids. Relevance is encoded by group_id matching.
This repository is part of the MLAIRE benchmark, submitted anonymously
to the NeurIPS 2026 Evaluations & Datasets Track. Authors and affiliations
are withheld for double-blind review.
Default top-k
Reported metrics in the paper use top-20.
Layout… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-submission-nips/mlaire-mlqa.
