BeIR/quora
Dataset Card for BEIR Benchmark quora is one of the datasets from the Duplicate Question Retrieval task within BEIR, measuring duplicate query retrieval for a given query. NOTE: ArguAna has queries also incorporated within the corpus, so you should remove the same query_id if present within the corpus during inference (implemented in BEIR) Dataset Summary BEIR is a heterogeneous benchmark built from 18 diverse datasets representing 9 information retrieval tasks.… See the full description on the dataset page: https://huggingface.co/datasets/BeIR/quora.
51.4k
Update README.md
Update README.md
Convert quora dataset to Parquet (#5)
Fix `license` metadata (#1)
add query and corpus together
Initial add
initial commit
