CoolFace
Datasetpublic

projecte-aina/viquiquad

ViquiQuAD: An Extractive QA Dataset for Catalan from Wikipedia Dataset Summary ViquiQuAD is an extractive Question Answering dataset for Catalan, built from the Catalan Wikipedia (Viquipèdia). 3,111 contexts extracted from 597 high-quality, original (non-translated) articles. For each context, 1 to 5 questions were created with their corresponding answers. Total: 15,153 question–answer pairs. This dataset can be used to fine-tune and evaluate extractive QA… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/viquiquad.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
0likes92downloads

projecte-aina/viquiquad · main · files are served by the source, never re-hosted here