CoolFace
Datasetpublic

allenai/multinews_sparse_max

This is a copy of the Multi-News dataset, except the input source documents of its test split have been replaced by a sparse retriever. The retrieval pipeline used: query: The summary field of each example corpus: The union of all documents in the train, validation and test splits retriever: BM25 via PyTerrier with default settings top-k strategy: "max", i.e. the number of documents retrieved, k, is set as the maximum number of documents seen across examples in this dataset, in this case… See the full description on the dataset page: https://huggingface.co/datasets/allenai/multinews_sparse_max.

sourceHugging Faceotherupdated 4y agoView on Hugging Face
0likes78downloads
settings

This repository belongs to allenai on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namemultinews_sparse_max
visibilitypublic
licenceother
gatedno
ownerallenai
Account settings
allenai/multinews_sparse_max · CoolFace