CoolFace
Datasetpublic

KennethEnevoldsen/spontanous-speech-qa

Spontanous speech QA This dataset contains QA pairs from the spontaneous speech subsection of the Danish Gigaword. The dataset is created from the DDSC dataset and filtered to only include QA pairs where the question is less than 20 tokens and the answer is at least 4 tokens long. To find out more about the creation see the accompanying script.

sourceHugging Faceupdated 3y agoView on Hugging Face
0likes16downloads
Dataset Card

Spontanous speech QA

This dataset contains QA pairs from the spontaneous speech subsection of the Danish Gigaword. The dataset is created from the DDSC dataset and filtered to only include QA pairs where the question is less than 20 tokens and the answer is at least 4 tokens long.

To find out more about the creation see the accompanying script.