CoolFace
Datasetpublic

Lala8383/ms-marco-qa-10k

MS MARCO QA Subset (10K) This is a subset of the MS MARCO v1.1 dataset by Microsoft, sampled for lightweight experimentation. Source Original dataset: microsoft/ms_marco (v1.1) Original paper: MS MARCO: A Human Generated MAchine Reading COmprehension Dataset Original authors: Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, Li Deng (Microsoft) What was changed Randomly sampled 10,000 examples from the train… See the full description on the dataset page: https://huggingface.co/datasets/Lala8383/ms-marco-qa-10k.

sourceHugging Faceotherupdated 6mo agoView on Hugging Face
0likes43downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face