datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
harmful-contents
Harmful-Contents Dataset
A multi-label image dataset for harmful-content classification across eight PEGI-aligned categories.The dataset consists of 5,153 rights-cleared images, split into train/validation/test sets and annotated with both binary labels and mask fields for controlled negative sampling.
Dataset Structure
Harmful-Contents/
csv/
train.csv
val.csv
test.csv
data/
train/*.jpg
val/*.jpg
test/*.jpg
Each CSV contains:
name,
alcohol… See the full description on the dataset page: https://huggingface.co/datasets/onullusoy/harmful-contents.germeval-2025-harmful-content-detection-training-dataset
GermEval 2025 Harmful Content Detection - Training Sets
(Call to Action • Attacks on Democratic Basic Order • Violence)
Author: Samuel Ruairí Bullard - University of Regensburg
Models: Model Zoo (Gradio Space)
Base model: LSX-UniWue/ModernGBERT_134M
Competition: GermEval 2025 Shared Task
Collection: GermEval 2025 Contribution CollectionabullardUR@GermEval Shared Task 2025 Submission
Dataset Summary
This repository republishes the training splits used… See the full description on the dataset page: https://huggingface.co/datasets/abullard1/germeval-2025-harmful-content-detection-training-dataset.harmful-content-datasetkorean-harmful-content-detectionkorean-harmful-content-detection-window3
