CoolFace
Datasetpublic

mteb/MLQARetrieval

MLQARetrieval An MTEB dataset Massive Text Embedding Benchmark MLQA (MultiLingual Question Answering) is a benchmark dataset for evaluating cross-lingual question answering performance. MLQA consists of over 5K extractive QA instances (12K in English) in SQuAD format in seven languages - English, Arabic, German, Spanish, Hindi, Vietnamese and Simplified Chinese. MLQA is highly parallel, with QA instances parallel between 4 different languages on average.… See the full description on the dataset page: https://huggingface.co/datasets/mteb/MLQARetrieval.

sourceHugging Facecc-by-sa-3.0updated 7mo agoView on Hugging Face
2likes1.7kdownloads

mteb/MLQARetrieval · main · files are served by the source, never re-hosted here