datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
reranker-scoresrerankers-and-lexical-similarities
Dataset Card for Re-ranker Evaluation Datasets
This repo contains the evaluation datasets used in the paper "Language Model Re-rankers are Fooled by Lexical Similarities" accepted to FEVER 2025.
Dataset Details
The datasets in this repo are based on the NQ, LitQA2 (from LAB-Bench) and DRUID datasets. More details on the datasets can be found in our paper.
Uses
Evaluate re-rankers.
Dataset Structure
We release the NQ, LitQA2 and DRUID… See the full description on the dataset page: https://huggingface.co/datasets/Lo/rerankers-and-lexical-similarities.
