fikeisjan/child-safety-alignment-dataset
Child-Safety Alignment Dataset Content warning. This dataset contains synthetic examples of harmful and manipulative language directed at minors. It exists to train and evaluate protective classifiers. Companion dataset to "Mind the Alignment Gap: Why General-Purpose Moderation Fails Children, and How a Child-Centric Taxonomy and Synthetic Data Close It" (WOAH 2026, EMNLP). Summary 79,193 fully synthetic child–AI interactions, labelled harmful vs. safe, spanning… See the full description on the dataset page: https://huggingface.co/datasets/fikeisjan/child-safety-alignment-dataset.
033
No commit history came back for main. The revision may not exist, or the source declined the request.
