datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llama-3.1-medprm-reward-training-set
Med-PRM-Reward (Version 1.0)
🚀 Med-PRM-Reward is among the first Process Reward Models (PRMs) specifically designed for the medical domain. Unlike conventional PRMs, it enhances its verification capabilities by integrating clinical knowledge through retrieval-augmented generation (RAG). Med-PRM-Reward demonstrates exceptional performance in scaling-test-time computation, particularly outperforming majority‐voting ensembles on complex medical reasoning tasks. Moreover, its… See the full description on the dataset page: https://huggingface.co/datasets/dmis-lab/llama-3.1-medprm-reward-training-set.llama-3.1-medprm-reward-test-set🚀 Med-PRM-Reward is among the first Process Reward Models (PRMs) specifically designed for the medical domain. Unlike conventional PRMs, it enhances its verification capabilities by integrating clinical knowledge through retrieval-augmented generation (RAG). Med-PRM-Reward demonstrates exceptional performance in scaling-test-time computation, particularly outperforming majority‐voting ensembles on complex medical reasoning tasks. Moreover, its scalability is not limited to… See the full description on the dataset page: https://huggingface.co/datasets/dmis-lab/llama-3.1-medprm-reward-test-set.ToxReason
ToxReason
🚀 Accepted at ACL 2026 Findings
ToxReason is a benchmark dataset for mechanistic chemical toxicity reasoning based on Adverse Outcome Pathways (AOPs).
The dataset is designed to evaluate whether large language models can generate biologically interpretable toxicity reasoning trajectories that connect molecular structures, Molecular Initiating Events (MIEs), pathway perturbations, and organ-level adverse outcomes.
Dataset Overview
ToxReason consists of… See the full description on the dataset page: https://huggingface.co/datasets/dmis-lab/ToxReason.TemporalHead
[ACL 2025] Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
This repository contains two separate subsets of data (configs):
Temporal: JSON files in Temporal that include temporal knowledge.
Invariant: JSON files in Invariant that describe time-invariant knowledge based on LRE.
Each subset has its own schema. By defining them as two configs in the YAML header above, Hugging Face’s Dataset Viewer will show “Temporal” and… See the full description on the dataset page: https://huggingface.co/datasets/dmis-lab/TemporalHead.
