datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hausa_common_voiceThis dataset is from the common voice corpus 7.0 using the Hausa dataset
hausa-ajami-blindspot-evalHausaHate
Evaluation Benchmark for Hausa Hate Speech Detection
We introduce the first expert annotated corpus of Facebook comments for Hausa hate speech detection.
The corpus titled HausaHate comprises 2,000 comments extracted from Western African Facebook pages and
manually annotated by three Hausa native speakers, who are also NLP experts.
The corpus was annotated using two different layers. We first labeled each comment according to a
binary classification: offensive versus… See the full description on the dataset page: https://huggingface.co/datasets/franciellevargas/HausaHate.abfall-im-haushalt-simuliert
Abfall im Haushalt (simuliert)
High-quality synthetic dataset for machine learning and data analysis.
📊 Dataset Overview
Rows: 500,000 (sample: 1,000)
Columns: 11
Quality Score: 99%
Format: CSV
Type: 100% Synthetic
🎯 Features
datum
kosten
abfall_menge
abfall_art
kosten_pro_kg
abfall_quartal
kosten_quartal
abfall_art_quartal
kosten_abfall_art_pro_kg
abfall_quartal_jahr
... and 1 more
💡 Use Cases
🤖 Machine learning model training
📊 Data… See the full description on the dataset page: https://huggingface.co/datasets/MarvHins/abfall-im-haushalt-simuliert.
