CoolFace
Datasetpublicgated

fikeisjan/child-safety-alignment-dataset

Child-Safety Alignment Dataset Content warning. This dataset contains synthetic examples of harmful and manipulative language directed at minors. It exists to train and evaluate protective classifiers. Companion dataset to "Mind the Alignment Gap: Why General-Purpose Moderation Fails Children, and How a Child-Centric Taxonomy and Synthetic Data Close It" (WOAH 2026, EMNLP). Summary 79,193 fully synthetic child–AI interactions, labelled harmful vs. safe, spanning… See the full description on the dataset page: https://huggingface.co/datasets/fikeisjan/child-safety-alignment-dataset.

sourceHugging Faceotherupdated 14d agoView on Hugging Face
0likes32downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.