datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CASIA_speech_emotion_recognitionspeech-emotion-recognition
Speech Emotion Recognition
Dataset comprises 30,000+ audio recordings featuring 4 distinct emotions: euphoria, joy, sadness, and surprise. This extensive collection is designed for research in emotion recognition, focusing on the nuances of emotional speech and the subtleties of speech signals as individuals vocally express their feelings.
By utilizing this dataset, researchers and developers can enhance their understanding of sentiment analysis and improve automatic speech… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/speech-emotion-recognition.ssi-speech-emotion-recognition
Dataset Card for SSI: Speech Emotion Recognition - Stapes AI
Dataset Details
Dataset Format for Audio Files
This is the format for the audio files in the dataset. We'll open-source the dataset soon.
Gender
M - Male
F - Female
Age Group
CH - Child (0-12)
TE - Teenager (13-19)
AD - Adult (20-60)
SE - Senior (60+)
UNK - Unknown
Utterance Type
SEN: Sentence
WOR: Word
PHR: Phrase
Sentence
DFA: "Don't Forget A… See the full description on the dataset page: https://huggingface.co/datasets/stapesai/ssi-speech-emotion-recognition.iemocap_emotion_recognition@article{busso2008iemocap,
title={IEMOCAP: Interactive emotional dyadic motion capture database},
author={Busso, Carlos and Bulut, Murtaza and Lee, Chi-Chun and Kazemzadeh, Abe and Mower, Emily and Kim, Samuel and Chang, Jeannette N and Lee, Sungbok and Narayanan, Shrikanth S},
journal={Language resources and evaluation},
volume={42},
pages={335--359},
year={2008},
publisher={Springer}
}
@article{wang2024audiobench,
title={AudioBench: A Universal Benchmark for Audio Large… See the full description on the dataset page: https://huggingface.co/datasets/AudioLLMs/iemocap_emotion_recognition.Moroccan-Arabic-Multimodal-Emotion-Recognition
MDER-MA — Moroccan Arabic Multimodal Emotion Recognition (TTS-aligned repackaging)
A repackaging of the MDER-MA dataset that pairs every audio clip with its Arabic (Moroccan dialect / Darija) transcript and ships speaker-disjoint train/validation/test splits.
Original dataset: Ouali, S. & El Garouani, S. (2025). MDER-MA: A multimodal dataset for emotion recognition in low-resource Moroccan Arabic language. Data in Brief. DOI: 10.1016/j.dib.2025.112005. Mendeley:… See the full description on the dataset page: https://huggingface.co/datasets/FatimahEmadEldin/Moroccan-Arabic-Multimodal-Emotion-Recognition.Emotion_RecognitionPromptTTS_Emotion_Recognition_8k
Dataset Card for "PromptTTS_Emotion_Recognition_8k"
More Information needed
EmotionRecognition_MultimodalEmotionlinesDataset
Dataset Card for "emotion_recognition_multimodal_emotionlines_dataset"
More Information needed
CASIA_speech_emotion_recognitionSpeech_Emotion_RecognitionPromptTTS_Emotion_Recognition
Dataset Card for "PromptTTS_Emotion_Recognition"
More Information needed
gender_emotion_recognitionEmoji-Grounded_Speech_Emotion_Recognition
Dataset Card for "Emoji-Grounded_Speech_Emotion_Recognition"
More Information needed
EmotionRecognition_MultimodalEmotionlinesDataset_TTSEmotionRecognition_MultimodalEmotionlinesDataset_TTSCASIA_speech_emotion_recognitionemotion_recognitionEmotionRecognition_MultimodalEmotionlinesDataset
