datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
kiki-bouba-audio-20k
🔊 Kiki–Bouba, Spoken Aloud (20k Global Responses)
Dataset Summary
This dataset is the audio companion to
Rapidata/psychology-association-kiki-bouba-etc.
In the original dataset, respondents read the question "Which one is called 'Kiki'?" as written text.
Here, respondents instead hear the word spoken aloud — the task shows the same two shapes
(a rounded blob and a spiky star) while a short audio clip of "kiki" or "bouba" plays as context.
The annotator UI… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/kiki-bouba-audio-20k.audio-emotion-detection-dataset
Audio Emotion Detection Dataset
Github: Audio Emotion Detection Dataset
Connect with me : Linkedin
Speech clips in English and Hindi annotated with emotion labels and ASR transcripts.
Audio is sourced from public YouTube videos and trimmed to approximately 60 seconds per clip.
Noise reduction is applied via noisereduce and silero-vad.
Emotions (5 classes)
Label
Description
angry
Aggressive, confrontational speech
calm… See the full description on the dataset page: https://huggingface.co/datasets/RapidOrc121/audio-emotion-detection-dataset.
