datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
msmarco_passage_ranking_corpusThis is the preprocessed data from msmarco passage(v1) ranking corpus.
MS MARCO: A human generated MAchine Reading COmprehension dataset SPayal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen,.
msmarco_passage_ranking_official_trainThis is the preprocessed training data from msmarco passage(v1) ranking corpus.
MS MARCO: A human generated MAchine Reading COmprehension dataset SPayal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen,.
E2Rank_ranking_datasetsmsmarco_passage_ranking_queriesThis is the preprocessed queries from msmarco passage(v1) ranking corpus.
MS MARCO: A human generated MAchine Reading COmprehension dataset SPayal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen,.
persian-response-ranking
PerSHOP Response Ranking
A Persian response-ranking benchmark derived from PerSHOP and used to evaluate lexical, semantic-embedding and LLM-based ranking methods. It includes the existing Random and Same-Domain configurations alongside a newly developed Domain-Lexical hard-negative configuration.
Each instance pairs a customer query with 5 candidate responses: 1 correct (gold) and 4 negatives. The task is to rank the gold response first.
The same 2,116 query–gold-response… See the full description on the dataset page: https://huggingface.co/datasets/blueharu/persian-response-ranking.TVR_rankinghuman_rankingsThis dataset supports the From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach paper.
Dataset Description
The dataset comprises human preference rankings of automatically generated video descriptions. For each video, users were asked to rank five descriptions based on criteria such as completeness and richness.
Citation
@misc{masala2025visionlanguagegraphevents,
title={From Vision To Language through Graph of… See the full description on the dataset page: https://huggingface.co/datasets/mihaimasala/human_rankings.msmarco_full_ranking_listvlm_jury_rankingsThis dataset supports the From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach paper.
Dataset Description
The dataset comprises VLMs preference rankings of automatically generated video descriptions. For each video, the VLMs were asked to rank six descriptions based on criteria such as completeness and richness.
Citation
@misc{masala2025visionlanguagegraphevents,
title={From Vision To Language through Graph of… See the full description on the dataset page: https://huggingface.co/datasets/mihaimasala/vlm_jury_rankings.rankingspeed-ranking
Speed Ranking
All 31 working dispatchAI models ranked by CPU inference speed.
🚀 dispatchAI
efficiency-ranking
Efficiency Ranking
Models ranked by tokens-per-second per MB of file size. Higher = more efficient.
🚀 dispatchAI
arcade-ranking-trainingrubric_concat_v0_v4_train_with_rankingmovie-ranking-200k
