datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
peerreview-bench
PeerReview Bench
CMU Paper Reviewer:https://prometheus-eval.github.io/cmu-paper-reviewer/
Repository:https://github.com/prometheus-eval/cmu-paper-reviewer
Paper:https://arxiv.org/abs/2605.20668
Point of Contact:seungone@kaist.ac.kr
Expert-annotated review items from scientific papers, organized for three
complementary evaluation tasks. All data in this dataset is intended
for evaluation, not training. All configs reference a shared, deduplicated
file store (submitted_papers)… See the full description on the dataset page: https://huggingface.co/datasets/prometheus-eval/peerreview-bench.ICLR_Peer_Reviews_2026ICLR_Peer_Reviews_2022ICLR_Peer_Reviews_2025ICLR_Peer_Reviews_2024firstpass-peer-review
FirstPass
FirstPass is a multi-domain, multi-round scientific peer-review dataset built from
Nature Communications transparent peer-review files. It contains 3,668 complete
peer-review dialogues across five scientific domains, with editorial outcome labels
derived from real revision-cycle outcomes.
The dataset was introduced alongside a fine-tuned model (Qwen2.5-7B-Instruct + LoRA)
that achieves 80.5% accuracy and F1-macro 78.2% on revision-cycle prediction
(STANDARD vs.… See the full description on the dataset page: https://huggingface.co/datasets/Prabhjotschugh/firstpass-peer-review.ICLR_Peer_Reviews
ICLR Peer Reviews Dataset
This dataset contains peer reviews from ICLR (International Conference on Learning Representations), processed and standardized.
Dataset Structure
title: Title of the paper
abstract: Abstract of the paper
full_text: Full text of the paper (if available)
review: The peer review text
source: Source conference/year (e.g., iclr2020)
review_src: Original source identifier
year: Year of the conference (extracted from source)
overall_score: Computed… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/ICLR_Peer_Reviews.ICLR_Peer_Reviews_2023iclr-2017-2020-peer-review-with-thinking-tracepeerreview-bench
PeerReview Bench
Expert-annotated review items from scientific papers, organized for three
complementary evaluation tasks. All data in this dataset is intended
for evaluation, not training. All configs reference a shared, deduplicated
file store (submitted_papers) via SHA256 content hashes.
Every config exposes a single eval split.
Configs
reviewer
For evaluating AI reviewers (models that generate reviews from a paper).
One row per paper.
Minimal fields:… See the full description on the dataset page: https://huggingface.co/datasets/nlile/peerreview-bench.ICLR_Peer_Reviews_2021BMC_Medicine_Peer_Reviews
BMC Medicine Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').
review_src:… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Medicine_Peer_Reviews.ICLR_Peer_Reviews_2019ICLR_Peer_Reviews_2020BMC_Cancer_Peer_Reviews
BMC Cancer Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').
review_src:… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Cancer_Peer_Reviews.BMC_Medical_Genomics_Peer_Reviews
BMC Medical Genomics Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Medical_Genomics_Peer_Reviews.ICLR_Peer_Reviews_2018BMC_Peer_ReviewsTPR_Peer_Reviews
Transparent Peer Review (TPR) Dataset
This dataset contains peer reviews from TPR, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('TPR').… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/TPR_Peer_Reviews.BMC_Womens_Health_Peer_Reviews
BMC Women's Health Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Womens_Health_Peer_Reviews.ICLR_Peer_Reviews_2017BMC_Gastroenterology_Peer_Reviews
BMC Gastroenterology Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Gastroenterology_Peer_Reviews.BMC_Surgery_Peer_Reviews
BMC Surgery Peer Reviews
This dataset contains peer reviews from BMC, standardized to match the format of the pawin205/PeerRT dataset.
Dataset Structure
Each record contains the following attributes:
relative_rank: Default value (0).
win_prob: Default value (0.0).
title: Title of the paper.
abstract: Abstract of the paper.
full_text: Full text of the paper (or review text if unavailable).
review: The peer review text.
source: Source of the data ('BMC').
review_src:… See the full description on the dataset page: https://huggingface.co/datasets/JerMa88/BMC_Surgery_Peer_Reviews.ICLR_Peer_Reviews_2016ICLR_Peer_Reviews_2014ICLR_Peer_Reviews_2013
