projecte-aina/viquiquad
ViquiQuAD: An Extractive QA Dataset for Catalan from Wikipedia Dataset Summary ViquiQuAD is an extractive Question Answering dataset for Catalan, built from the Catalan Wikipedia (Viquipèdia). 3,111 contexts extracted from 597 high-quality, original (non-translated) articles. For each context, 1 to 5 questions were created with their corresponding answers. Total: 15,153 question–answer pairs. This dataset can be used to fine-tune and evaluate extractive QA… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/viquiquad.
Update README.md (#2)
Dataset viewer problem: Final modifications
New train data re-structured
Train data re-structured
Dataset viewer problem solved
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
fix language tag
Fix `license` metadata (#1)
Change title of dataset card
Update dataset card
Fix dataset tags
Fix style
Remove unnecessary extract
Remove unnecessary builder config
Fix TypeError
updated README
Update README.md
update
Update viquiquad.py
upload dataset
initial commit
