datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
coqa-flatcoqar-clarifications
CoQAR Clarifications
This dataset pairs 1,000 CoQAR development questions with their original stories and stories damaged by sentence deletion. Each of the resulting 2,000 inputs has five sampled model clarifications. Two configurations reuse the same generated additions and differ only in where those additions are placed.
Configuration
Rows in dev
Clarifications per row
Placement
appended
2,000
5
At the end of the input story
inserted
2,000
5
At the deleted passage… See the full description on the dataset page: https://huggingface.co/datasets/rvashurin/coqar-clarifications.coqatokenized_coqa_size356coqa_ask4conftokenized_coqa
