fikeisjan/child-safety-alignment-dataset
Child-Safety Alignment Dataset Content warning. This dataset contains synthetic examples of harmful and manipulative language directed at minors. It exists to train and evaluate protective classifiers. Companion dataset to "Mind the Alignment Gap: Why General-Purpose Moderation Fails Children, and How a Child-Centric Taxonomy and Synthetic Data Close It" (WOAH 2026, EMNLP). Summary 79,193 fully synthetic child–AI interactions, labelled harmful vs. safe, spanning… See the full description on the dataset page: https://huggingface.co/datasets/fikeisjan/child-safety-alignment-dataset.
032
No card is published for this repository, or it could not be fetched from Hugging Face right now.
