ipda
Datasets
All datasets matching “ipda”ipda-judge-adaptation-grpo
IPDA Judge Adaptation GRPO Dataset
Training data for judge adaptation in competitive debate. Contains GRPO preference sets for adapting debate speech generation to different judge profiles.
Dataset Description
This dataset enables training LLMs to adapt their debate arguments based on judge characteristics:
Depth Adaptation: Adapting explanation complexity to judge expertise level (debate experience + domain knowledge)
Bias Adaptation: Adapting argument framing to judge… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-judge-adaptation-grpo.ipda-golden-samples
IPDA Golden Samples (2AR + 1AR)
Golden samples for fine-tuning debate models on affirmative rebuttal speeches in IPDA format.
Dataset Description
874 high-quality samples for SFT training:
447 2AR (Second Affirmative Rebuttal)
427 1AR (First Affirmative Rebuttal)
Dataset Sources
Source
Count
Description
iter2_group_c
832
High-scoring (>=0.75) samples from GRPO iteration 2
augmented_claude-opus-4.5
20
Augmented debates generated by Claude Opus 4.5… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-golden-samples.ipda-2ar-golden-samples
IPDA Golden Samples (2AR + 1AR)
Golden samples for fine-tuning debate models on affirmative rebuttal speeches in IPDA format.
Dataset Description
422 high-quality samples for SFT training:
260 2AR (Second Affirmative Rebuttal)
162 1AR (First Affirmative Rebuttal)
Dataset Sources
Model
2AR
1AR
Total
Claude Opus 4.5
100
50
150
GPT-5.2
100
50
150
Claude Sonnet
10
10
20
Claude Haiku
9
9
18
Qwen-ft (debate model)
19
19
38
Qwen-base
16
18
34… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-2ar-golden-samples.ipda-grpo-training-data
IPDA GRPO Training Data
Training data for GRPO (Group Relative Policy Optimization) on IPDA debate tasks.
Dataset Description
Contains scored debate speech samples used for GRPO training iterations. Each sample includes:
Input prompt (debate context)
Generated response (speech)
Rubric scores from debate judge
Log probabilities for policy optimization
Files
File
Description
Samples
group_c_grpo.parquet
Group C (warrant/clash) training data
~3K… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-grpo-training-data.ipda-grpo-nc-branched-v2ipda-sentence-selection-data
IPDA Sentence Selection Training Dataset
Training data for sentence-level claim selection in competitive debate. This dataset teaches models to select the most impactful claims to address during rebuttal speeches.
Dataset Structure
Files
File
Size
Description
sentence_selection_dataset.json
23MB
Full sentence selection dataset
sentence_dpo_format_consistent.json
4.4MB
DPO preference pairs (consistent format)
sentence_sft_train_v2.json
13MB
SFT… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-sentence-selection-data.
