hakari-bench/NanoR2MED
NanoR2MED This dataset is a Nano-style retrieval dataset for HAKARI-bench. NanoR2MED contains 8 Nano retrieval splits derived from R2MED. Each split keeps up to 200 eligible queries and up to 10000 corpus documents, with exact duplicate query and document text removed where the generator records that policy. Usage from datasets import load_dataset dataset_id = "hakari-bench/NanoR2MED" split = "NanoR2MEDBioinformatics" queries = load_dataset(dataset_id, "queries"… See the full description on the dataset page: https://huggingface.co/datasets/hakari-bench/NanoR2MED.
Update dataset README for reranking_hybrid candidates
Update dataset README for reranking_hybrid candidates
Add reranking_hybrid metadata
Update reranking_hybrid for reranking_hybrid candidates
Update harrier_oss_v1_270m for reranking_hybrid candidates
Update bm25 for reranking_hybrid candidates
Update qrels for reranking_hybrid candidates
Update queries for reranking_hybrid candidates
Update corpus for reranking_hybrid candidates
Copy dataset from hotchpotch/NanoR2MED
initial commit
