projecte-aina/vilaquad
Dataset Card for VilaQuAD Dataset Summary VilaQuAD, An extractive QA dataset for Catalan, from VilaWeb newswire text. This dataset contains 2095 of Catalan language news articles along with 1 to 5 questions referring to each fragment (or context). VilaQuad articles are extracted from the daily VilaWeb and used under CC-BY-NC-SA-ND licence. This dataset can be used to build extractive-QA and Language Models. Supported Tasks and Leaderboards… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/vilaquad.
This repository belongs to projecte-aina on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
