allenai/wcep_sparse_max
This is a copy of the WCEP-10 dataset, except the input source documents of its test split have been replaced by a sparse retriever. The retrieval pipeline used: query: The summary field of each example corpus: The union of all documents in the train, validation and test splits retriever: BM25 via PyTerrier with default settings top-k strategy: "max", i.e. the number of documents retrieved, k, is set as the maximum number of documents seen across examples in this dataset, in this case k==10… See the full description on the dataset page: https://huggingface.co/datasets/allenai/wcep_sparse_max.
Update README.md
Update README.md
Update README.md
Update README.md
Create README.md
Upload dataset_infos.json with huggingface_hub
Upload dataset_infos.json with huggingface_hub
Upload data/test-00000-of-00001-6133f005add30cdb.parquet with huggingface_hub
Upload data/validation-00000-of-00001-6d9e90279dc820aa.parquet with huggingface_hub
Upload data/train-00000-of-00001-af48311c4f368aff.parquet with huggingface_hub
initial commit
