CoolFace
20 results

ipda

debaterhub /ipda-judge-adaptation-grpo IPDA Judge Adaptation GRPO Dataset Training data for judge adaptation in competitive debate. Contains GRPO preference sets for adapting debate speech generation to different judge profiles. Dataset Description This dataset enables training LLMs to adapt their debate arguments based on judge characteristics: Depth Adaptation: Adapting explanation complexity to judge expertise level (debate experience + domain knowledge) Bias Adaptation: Adapting argument framing to judge… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-judge-adaptation-grpo.texttext-generationn<1K0 likes49 downloads8mo agoHugging Facedgonier /ipda-golden-samples IPDA Golden Samples (2AR + 1AR) Golden samples for fine-tuning debate models on affirmative rebuttal speeches in IPDA format. Dataset Description 874 high-quality samples for SFT training: 447 2AR (Second Affirmative Rebuttal) 427 1AR (First Affirmative Rebuttal) Dataset Sources Source Count Description iter2_group_c 832 High-scoring (>=0.75) samples from GRPO iteration 2 augmented_claude-opus-4.5 20 Augmented debates generated by Claude Opus 4.5… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-golden-samples.texttext-generationn<1K0 likes38 downloads8mo agoHugging Facedgonier /ipda-2ar-golden-samples IPDA Golden Samples (2AR + 1AR) Golden samples for fine-tuning debate models on affirmative rebuttal speeches in IPDA format. Dataset Description 422 high-quality samples for SFT training: 260 2AR (Second Affirmative Rebuttal) 162 1AR (First Affirmative Rebuttal) Dataset Sources Model 2AR 1AR Total Claude Opus 4.5 100 50 150 GPT-5.2 100 50 150 Claude Sonnet 10 10 20 Claude Haiku 9 9 18 Qwen-ft (debate model) 19 19 38 Qwen-base 16 18 34… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-2ar-golden-samples.texttext-generationn<1K0 likes29 downloads8mo agoHugging Facedgonier /ipda-grpo-training-data IPDA GRPO Training Data Training data for GRPO (Group Relative Policy Optimization) on IPDA debate tasks. Dataset Description Contains scored debate speech samples used for GRPO training iterations. Each sample includes: Input prompt (debate context) Generated response (speech) Rubric scores from debate judge Log probabilities for policy optimization Files File Description Samples group_c_grpo.parquet Group C (warrant/clash) training data ~3K… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-grpo-training-data.tabulartext-generation1K<n<10K0 likes26 downloads8mo agoHugging Facedgonier /ipda-grpo-nc-branched-v20 likes24 downloads7mo agoHugging Facedebaterhub /ipda-sentence-selection-data IPDA Sentence Selection Training Dataset Training data for sentence-level claim selection in competitive debate. This dataset teaches models to select the most impactful claims to address during rebuttal speeches. Dataset Structure Files File Size Description sentence_selection_dataset.json 23MB Full sentence selection dataset sentence_dpo_format_consistent.json 4.4MB DPO preference pairs (consistent format) sentence_sft_train_v2.json 13MB SFT… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/ipda-sentence-selection-data.text-generation10K<n<100K0 likes20 downloads9mo agoHugging Face