datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
WAMA-West-African-Marketing-Annotation-Dataset
license: cc-by-4.0
task_categories:
text-classification
text-generation
language:
en
pcm
tags:
marketing
west-africa
nigeria
ghana
consumer-psychology
trust-signals
nigerian-english
ghanaian-english
nigerian-pidgin
cultural-bias
ai-alignment
annotation
fintech
brand-strategy
underrepresented
africa
size_categories:
n<1K
WAMA West African Marketing Annotation Dataset
Version: 1.0Entries: 300Countries: Nigeria (165 entries), Ghana (135 entries)Created by: Deborah John Digital… See the full description on the dataset page: https://huggingface.co/datasets/Debbyjaye001/WAMA-West-African-Marketing-Annotation-Dataset.WAMP-Router-Intent-Dataset
WAMP Router Intent Dataset
This dataset was created for training the WAMP-proxy semantic router. It contains user queries in Russian and English across various domains (Technical, Medicine, Art, Philosophy, etc.) classified into three intent categories.
Dataset Structure
The dataset consists of two columns:
text: The user query string.
label: The integer class ID.
label_text: Human-readable class name.
Labels
0 (Summary): General requests for recaps, TL;DR… See the full description on the dataset page: https://huggingface.co/datasets/naranor/WAMP-Router-Intent-Dataset.german-spoon-language
Spoon Language
This is a dataset of random german sentences from Tatoeba (accessed 23.07.25) mapped to their "Löffelsprache" / spoon language version.
The corresponding GitHub repository can be found here.
