din0s/msmarco-nlgen
Dataset Card for MSMARCO - Natural Language Generation Task Dataset Summary The original focus of MSMARCO was to provide a corpus for training and testing systems which given a real domain user query systems would then provide the most likley candidate answer and do so in language which was natural and conversational. All questions have been generated from real anonymized Bing user queries which grounds the dataset in a real world problem and can provide… See the full description on the dataset page: https://huggingface.co/datasets/din0s/msmarco-nlgen.
646
