CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01seantw /DEBATE_LLM DEBATE Benchmark This repository contains CSV files from the DEBATE project: large-scale human conversation experiments organized around controversial and opinion-based topics. The data consists of multi-round conversations between human participants discussing political, social, and belief-related topics, following the protocol described in: Chuang, Y.-S., Tu, R., Dai, C., Vasani, S., Li, Y., Yao, B., Tessler, M. H., Yang, S., Shah, D., Hawkins, R., Hu, J., & Rogers, T. T. (2026).… See the full description on the dataset page: https://huggingface.co/datasets/seantw/DEBATE_LLM.tabular100K<n<1M4 likes368 downloads5mo agoHugging Face02Hellisotherpeople /DebateSum DebateSum Corresponding code repo for the upcoming paper at ARGMIN 2020: "DebateSum: A large-scale argument mining and summarization dataset" Arxiv pre-print available here: https://arxiv.org/abs/2011.07251 Check out the presentation date and time here: https://argmining2020.i3s.unice.fr/node/9 Full paper as presented by the ACL is here: https://www.aclweb.org/anthology/2020.argmining-1.1/ Video of presentation at COLING 2020:… See the full description on the dataset page: https://huggingface.co/datasets/Hellisotherpeople/DebateSum.tabularquestion-answering100K<n<1M21 likes291 downloads4y agoHugging Face03kokhayas /english-debate-motions-utdsEnglish Debate Motions gathered by University of Tokyo Debate Society @misc{english-debate-motions-utds, title={english-debate-motions-utds}, author={members of the University of Tokyo Debate Society}, year={2022}, } tabular10K<n<100K3 likes135 downloads4y agoHugging Face04democratic-commons /steuer_debateThe Steuer Debate consultation is a german citizen participation project on fair taxes and finances held in 2025. Data This dataset contains three subsets: proposals contains the written propositions (in german). the topic column contains the LLM-generated and human-validated topics (clusters) used during the official analysis of the consultation. Each proposal has a unique id proposal_id. votes contains the votes of users on propositions. Each user has a unique id user_id and… See the full description on the dataset page: https://huggingface.co/datasets/democratic-commons/steuer_debate.tabular10K<n<100K0 likes68 downloads12d agoHugging Face05frasalvi /debategpt Dataset Card for DebateGPT The DebateGPT dataset contains debates between humans and GPT-4, along with sociodemographic information about human participants and their agreement scores before and after the debates. This dataset was created for research on measuring the persuasiveness of language models and the impact of personalization, as described in this paper: On the Conversational Persuasiveness of GPT-4. Dataset Details The dataset consists of a CSV file with the… See the full description on the dataset page: https://huggingface.co/datasets/frasalvi/debategpt.tabularn<1K1 likes62 downloads1y agoHugging Face06NLP-Debater-Project /IBM-Debater-ArgKPtabular10K<n<100K3 likes55 downloads10mo agoHugging Face07PixelBeacon7 /dirty-debate-977cdf dirty-debate-977cdf Synthetic sensors test data: 45 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/PixelBeacon7/dirty-debate-977cdf.tabularn<1K0 likes31 downloads12d agoHugging Face08barissozudogru /belnap-debate-corpus Belnap Real-Debate Corpus (110 Propositions) What this is 110 contested propositions across 31 domains, used as a real-debate evaluation corpus for paraconsistent (Belnap-Dunn) debate aggregation. Each row is a single proposition sourced from established controversy databases. Distribution Source Count Kialo 52 ProCon.org 40 Classical philosophy 18 Controversy level Count high 79 medium 25 low 6 31 domains… See the full description on the dataset page: https://huggingface.co/datasets/barissozudogru/belnap-debate-corpus.texttext-classificationn<1K0 likes28 downloads4mo agoHugging Face09Malek-Messaoudi /IBM-Debater-Datasettabular10K<n<100K0 likes27 downloads9mo agoHugging Face10debatellm /DEBATE DEBATE Benchmark This repository contains CSV files from the DEBATE project: large-scale human conversation experiments organized around controversial and opinion-based topics. The data consists of multi-round conversations between human participants discussing political, social, and belief-related topics, following the protocol described in: Directory Structure . ├── raw/ # Raw exports │ ├── depth/ # Topic Set 1: Depth topics… See the full description on the dataset page: https://huggingface.co/datasets/debatellm/DEBATE.tabular100K<n<1M0 likes26 downloads5mo agoHugging Face11frasalvi /debategpt_topic_scoresThis dataset contains human annotator scores for the topic used in debates described in the paper: On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial . tabularn<1K0 likes21 downloads2y agoHugging Face12debate-land /2023-paradigms Dataset Card for Dataset Name Dataset Summary This is a list of approximately 4,700 judge "paradigms" sourced from Tabroom as a part of Debate Land's scraping. It comes from the 2023 National Circuit tournaments hosted on the website. Languages English. Dataset Structure Data Fields Field Description Paradigm The raw text of the judge's paradigm. FlowType What the judge's flowing behavior was interpreted as. (Flow, Flay, Lay… See the full description on the dataset page: https://huggingface.co/datasets/debate-land/2023-paradigms.textn<1K2 likes20 downloads3y agoHugging Face13DebateLabKIT /role-conflict-benchfrom: https://github.com/ddindidu/RoleConflictBench tabular10K<n<100K0 likes19 downloads8mo agoHugging Face14yassine-mhirsi /IBM-Debater-ArgKPtabular10K<n<100K0 likes16 downloads10mo agoHugging Face15momererkoc /turkish-debate-topics-datasettextn<1K1 likes14 downloads1y agoHugging Face16raphaaal /schopenhauer-debateFine-tuning dataset for creating an argumentative agent, following Schopenhauer's stratagems. Credits (GitHub): @basileplus, @vdeva, @mcosson , @yanisgomes, @raphaaal This dataset contains 1,000 conversations, generated synthetically using Mistral-Large. Each conversation starts with a claim from an Opponent and contains between 1 and 5 tweets debating this claim. Topic: the Silicon Valley Bank run debate on Twitter. Input: the beggining of a conversation between two Users (Opponent and You)… See the full description on the dataset page: https://huggingface.co/datasets/raphaaal/schopenhauer-debate.texttable-question-answeringn<1K5 likes13 downloads2y agoHugging Face17DebateLabKIT /billsum_train Mirror of billsum train split Mirror with parquet files on hub, as downloading billsum data files from Google drive causes errors in distributed training. text10K<n<100K0 likes12 downloads4y agoHugging Face18debate-land /prog-paradigmstextn<1K0 likes12 downloads3y agoHugging Face19debate-land /flow-paradigmstextn<1K0 likes11 downloads3y agoHugging Face20as-cle-bert /DebateLLMstextn<1K4 likes11 downloads2y agoHugging Face21dgonier /debate-llmtext100K<n<1M0 likes7 downloads2y agoHugging Face22mexca /comptext-workshop-debate-fulltabular1K<n<10K1 likes7 downloads2y agoHugging Face23LIXXX /ibm_debatetext1K<n<10K0 likes3 downloads4y agoHugging Face24mdudek /polish-presidential-debategatedaudion<1K0 likes1 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.