quicktensor/blockrank-msmarco-train-10p
BlockRank MS MARCO Training Data (10% Sample) Dataset Description A 10% sample of MS MARCO passage ranking data formatted for training in-context ranking LLMs. This dataset is used in the training of the BlockRank project: Scalable In-context Ranking with Generative Models. Format: JSONL (in-context ranking format) Size: 50k training examples (10% sample) Documents per query: 30-50 candidates (mix of positives and hard negatives) Source Original:… See the full description on the dataset page: https://huggingface.co/datasets/quicktensor/blockrank-msmarco-train-10p.
This repository belongs to quicktensor on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
