datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ipda-grpo-training-data
IPDA GRPO Training Data
Training data for GRPO (Group Relative Policy Optimization) on IPDA debate tasks.
Dataset Description
Contains scored debate speech samples used for GRPO training iterations. Each sample includes:
Input prompt (debate context)
Generated response (speech)
Rubric scores from debate judge
Log probabilities for policy optimization
Files
File
Description
Samples
group_c_grpo.parquet
Group C (warrant/clash) training data
~3K… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-grpo-training-data.ipda-phase5-v2
IPDA Debate Training Data - Phase 5 Iteration V2
Training data for IPDA (International Public Debate Association) debate AI model.
Dataset Description
This dataset contains per-call training examples extracted from full debate simulations, with quality scores assigned by a DSPy-based evaluation pipeline.
Pipeline Overview
Full Debate Generation: Complete IPDA debates generated using a DSPy pipeline with:
Multi-hop research via Tavily API
Structured speech… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-phase5-v2.ipda-judge-adaptation-data
IPDA Judge Adaptation Training Dataset
Training data for judge adaptation in competitive debate. This dataset teaches models to adapt their debate output based on judge characteristics.
Dataset Structure
Files
File
Description
Pairs
depth_iter1_train.json
Depth adaptation iteration 1 (lay vs expert judges)
75
depth_iter2_train.json
Depth adaptation iteration 2 (different topics)
75
bias_train.json
Bias adaptation (ideological, procedural… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-judge-adaptation-data.ipda_grpo_multi_trial_thinking_tactics
Debate Multi-Trial GRPO Test Data (with Thinking Frameworks)
TEST DATASET - Single debate for review before scaling.
Training data for offline GRPO (Group Relative Policy Optimization) on IPDA debate generation,
with integrated thinking framework injection.
What's New: Thinking Frameworks
Each prompt includes structured thinking instructions (mnemonics) that guide the model's reasoning:
Call Type
Mnemonic
Purpose
TACTIC_SELECT
JAM
Judge-Attack-Momentum Analysis… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda_grpo_multi_trial_thinking_tactics.ipda-grpo-dataset-iter3-feb-12
IPDA GRPO Dataset — Iteration 3 (Feb 12, 2026)
GRPO (Group Relative Policy Optimization) training dataset for IPDA (International Public Debate Association) debate speech generation.
Dataset Structure
2,988 unique prompts | 11,425 scored trials | Score avg: 0.700 (0-1 scale)
Each row represents a unique debate pipeline prompt with up to 6 trial responses:
Column
Description
prompt_hash
SHA256[:16] of prompt text
prompt
Full pipeline prompt
speech_type
AC… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-grpo-dataset-iter3-feb-12.
