CoolFace
Datasetpublic

ucsahin/TR-Extractive-QA-82K

The dataset consists of nearly 82K {Context, Question, Answer} triplets in Turkish. Since most of the answers are only a few words and taken directly from the provided context, it can be better used in in finetuning encoder-only models like BERT for extractive question answering or embedding models for retrieval. The dataset is a filtered and combined version of multiple Turkish QA-based datasets. Please use ucsahin/TR-Extractive-QA-5K for more detailed and sampled version of this dataset.

sourceHugging Faceupdated 2y agoView on Hugging Face
6likes74downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ucsahin/TR-Extractive-QA-82K · CoolFace