datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
blockrank-beir-evals
ICR-BEIR-Evals: In-Context Ranking Evaluation Dataset
Dataset Description
ICR-BEIR-Evals is a curated evaluation dataset for In-Context Ranking (ICR) models, derived from the BEIR benchmark. This dataset is specifically designed to evaluate the effectiveness of generative language models on document ranking tasks where queries and candidate documents are provided in-context.
The dataset contains 28,759 queries across 11 diverse BEIR datasets, with each query paired with… See the full description on the dataset page: https://huggingface.co/datasets/quicktensor/blockrank-beir-evals.blockrank-msmarco-train-10p
BlockRank MS MARCO Training Data (10% Sample)
Dataset Description
A 10% sample of MS MARCO passage ranking data formatted for training in-context ranking LLMs. This dataset is used in the training of the BlockRank project: Scalable In-context Ranking with Generative Models.
Format: JSONL (in-context ranking format)
Size: 50k training examples (10% sample)
Documents per query: 30-50 candidates (mix of positives and hard negatives)
Source
Original: MS… See the full description on the dataset page: https://huggingface.co/datasets/quicktensor/blockrank-msmarco-train-10p.blockrank-data
