CoolFace
Datasetpublic

quicktensor/blockrank-msmarco-train-10p

BlockRank MS MARCO Training Data (10% Sample) Dataset Description A 10% sample of MS MARCO passage ranking data formatted for training in-context ranking LLMs. This dataset is used in the training of the BlockRank project: Scalable In-context Ranking with Generative Models. Format: JSONL (in-context ranking format) Size: 50k training examples (10% sample) Documents per query: 30-50 candidates (mix of positives and hard negatives) Source Original:… See the full description on the dataset page: https://huggingface.co/datasets/quicktensor/blockrank-msmarco-train-10p.

sourceHugging Faceapache-2.0updated 11mo agoView on Hugging Face
1likes18downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face