Harikrishna-Srinivasan/Hate-Speech
Hate Speech This dataset contains 25,001 English social media sentences labeled for binary hate speech classification. text - input sentence label - 0 (Not Hate), 1 (Hate) Source This dataset is a cleaned and consolidated version of: Waal, B. — Hate Speech Detection – Curated Datasethttps://www.kaggle.com/datasets/waalbannyantudre/hate-speech-detection-curated-dataset Morjaria, M. — Hate Speech and Offensive Language… See the full description on the dataset page: https://huggingface.co/datasets/Harikrishna-Srinivasan/Hate-Speech.
Hate Speech
This dataset contains 25,001 English social media sentences labeled for binary hate speech classification.
text- input sentencelabel-0(Not Hate),1(Hate)
Source
This dataset is a cleaned and consolidated version of:
- Waal, B. — Hate Speech Detection – Curated Dataset https://www.kaggle.com/datasets/waalbannyantudre/hate-speech-detection-curated-dataset
- Morjaria, M. — Hate Speech and Offensive Language Dataset https://www.kaggle.com/datasets/mrmorj/hate-speech-and-offensive-language-dataset
Additional preprocessing performed by Harikrishna Srinivasan (2026).
Citation
@misc{srinivasan2026hatespeech,
author = {Harikrishna Srinivasan},
title = {Hate-Speech Dataset (Refined and Cleaned Version)},
year = {2026},
publisher = {Hugging Face Datasets},
howpublished = {\url{https://huggingface.co/datasets/Harikrishna-Srinivasan/Hate-Speech}}
}
