CoolFace
Datasetpublic

ucsahin/TR-Extractive-QA-82K

The dataset consists of nearly 82K {Context, Question, Answer} triplets in Turkish. Since most of the answers are only a few words and taken directly from the provided context, it can be better used in in finetuning encoder-only models like BERT for extractive question answering or embedding models for retrieval. The dataset is a filtered and combined version of multiple Turkish QA-based datasets. Please use ucsahin/TR-Extractive-QA-5K for more detailed and sampled version of this dataset.

sourceHugging Faceupdated 2y agoView on Hugging Face
6likes70downloads
5 commits on main
c539d512y ago

Update README.md

ucsahin
c5f2cbf2y ago

Update README.md

ucsahin
e7c87a92y ago

Update README.md

ucsahin
f2f25d32y ago

Upload dataset

ucsahin
04d38c52y ago

initial commit

ucsahin