nlpatunt/D_persuade_2
Persuade_2 The PERSUADE 2.0 corpus (Persuasive Essays for Rating, Selecting, and Understanding Argumentative and Discourse Elements) contains over 25,000 argumentative essays written by 6th–12th grade students in the United States, covering 15 distinct prompts across two writing tasks: independent and source-based writing. The corpus also provides detailed individual and demographic information for each writer. This is the train, test, and validation split of the dataset.… See the full description on the dataset page: https://huggingface.co/datasets/nlpatunt/D_persuade_2.
Persuade_2
The PERSUADE 2.0 corpus (Persuasive Essays for Rating, Selecting, and Understanding Argumentative and Discourse Elements) contains over 25,000 argumentative essays written by 6th–12th grade students in the United States, covering 15 distinct prompts across two writing tasks: independent and source-based writing. The corpus also provides detailed individual and demographic information for each writer.
This is the train, test, and validation split of the dataset. Ground truth labels have been removed to prevent leakage during evaluation.This version is prepared for use with the S-GRADES benchmark.
Original Dataset
Citation
If you use this dataset, please cite the original authors:
@article{crossley2024persuade,
title={A large-scale corpus for assessing written argumentation: PERSUADE 2.0},
author={Crossley, Scott A. and Tian, Yanwen and Baffour, Paul and Franklin,
Ashley and Benner, Matthew and Boser, Ulrich},
journal={Assessing Writing},
volume={61},
pages={100865},
year={2024},
issn={1075-2935},
doi={10.1016/j.asw.2024.100865}
}If used as part of S-GRADES, also cite:
@inproceedings{seuti2026sgrades,
title={S-GRADES: Studying Generalization of Student Response Assessments in Diverse Evaluative Settings},
author={Seuti, Tasfia and Ray Choudhury, Sagnik},
booktitle={Proceedings of the 15th International Conference on Language Resources and Evaluation (LREC 2026)},
year={2026}
}