datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CASIA_speech_emotion_recognitionvideo-emotion-recognition-dataset
Video Dataset of Various Emotions for Recognition Tasks
Dataset comprises 1,000+ videos featuring 11 facial emotions and 15 inner emotions expressed by individuals from diverse backgrounds, including various races, genders, and ages. It is designed for emotion recognition research, focusing on emotion detection and emotion classification tasks.
By utilizing this dataset, researchers can explore advanced emotion analysis techniques and develop robust recognition models that can… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/video-emotion-recognition-dataset.speech-emotion-recognition
Speech Emotion Recognition
Dataset comprises 30,000+ audio recordings featuring 4 distinct emotions: euphoria, joy, sadness, and surprise. This extensive collection is designed for research in emotion recognition, focusing on the nuances of emotional speech and the subtleties of speech signals as individuals vocally express their feelings.
By utilizing this dataset, researchers and developers can enhance their understanding of sentiment analysis and improve automatic speech… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/speech-emotion-recognition.ssi-speech-emotion-recognition
Dataset Card for SSI: Speech Emotion Recognition - Stapes AI
Dataset Details
Dataset Format for Audio Files
This is the format for the audio files in the dataset. We'll open-source the dataset soon.
Gender
M - Male
F - Female
Age Group
CH - Child (0-12)
TE - Teenager (13-19)
AD - Adult (20-60)
SE - Senior (60+)
UNK - Unknown
Utterance Type
SEN: Sentence
WOR: Word
PHR: Phrase
Sentence
DFA: "Don't Forget A… See the full description on the dataset page: https://huggingface.co/datasets/stapesai/ssi-speech-emotion-recognition.iemocap_emotion_recognition@article{busso2008iemocap,
title={IEMOCAP: Interactive emotional dyadic motion capture database},
author={Busso, Carlos and Bulut, Murtaza and Lee, Chi-Chun and Kazemzadeh, Abe and Mower, Emily and Kim, Samuel and Chang, Jeannette N and Lee, Sungbok and Narayanan, Shrikanth S},
journal={Language resources and evaluation},
volume={42},
pages={335--359},
year={2008},
publisher={Springer}
}
@article{wang2024audiobench,
title={AudioBench: A Universal Benchmark for Audio Large… See the full description on the dataset page: https://huggingface.co/datasets/AudioLLMs/iemocap_emotion_recognition.facial-emotion-recognition-datasetThe dataset consists of images capturing people displaying 7 distinct emotions
(anger, contempt, disgust, fear, happiness, sadness and surprise).
Each image in the dataset represents one of these specific emotions,
enabling researchers and machine learning practitioners to study and develop
models for emotion recognition and analysis.
The images encompass a diverse range of individuals, including different
genders, ethnicities, and age groups*. The dataset aims to provide
a comprehensive representation of human emotions, allowing for a wide range of
use cases.Human-Face_Images_for_Emotion_RecognitionMoroccan-Arabic-Multimodal-Emotion-Recognition
MDER-MA — Moroccan Arabic Multimodal Emotion Recognition (TTS-aligned repackaging)
A repackaging of the MDER-MA dataset that pairs every audio clip with its Arabic (Moroccan dialect / Darija) transcript and ships speaker-disjoint train/validation/test splits.
Original dataset: Ouali, S. & El Garouani, S. (2025). MDER-MA: A multimodal dataset for emotion recognition in low-resource Moroccan Arabic language. Data in Brief. DOI: 10.1016/j.dib.2025.112005. Mendeley:… See the full description on the dataset page: https://huggingface.co/datasets/FatimahEmadEldin/Moroccan-Arabic-Multimodal-Emotion-Recognition.Emotion_Recognitionexpress-emotion-recognition
EXPRESS Dataset
Overview
This is the EXPRESS dataset from the paper:
Fluent but Unfeeling: The Emotional Blind Spots of Language Models
EXPRESS (EXperiences and PRocessed Emotions in Self-disclosure Stories) is a benchmark dataset for evaluating fine-grained emotion recognition in language models. It contains 33,679 naturally occurring Reddit-based human experiences paired with self-disclosed emotion labels.
EXPRESS uses emotions explicitly disclosed by the… See the full description on the dataset page: https://huggingface.co/datasets/bangzhao/express-emotion-recognition.speech-emotion-recognition-datasetThe audio dataset consists of a collection of texts spoken with four distinct
emotions. These texts are spoken in English and represent four different
emotional states: **euphoria, joy, sadness and surprise**.
Each audio clip captures the tone, intonation, and nuances of speech as
individuals convey their emotions through their voice.
The dataset includes a diverse range of speakers, ensuring variability in age,
gender, and cultural backgrounds*, allowing for a more comprehensive
representation of the emotional spectrum.
The dataset is labeled and organized based on the emotion expressed in each
audio sample, making it a valuable resource for emotion recognition and
analysis. Researchers and developers can utilize this dataset to train and
evaluate machine learning models and algorithms, aiming to accurately
recognize and classify emotions in speech.PromptTTS_Emotion_Recognition_8k
Dataset Card for "PromptTTS_Emotion_Recognition_8k"
More Information needed
EmotionRecognition_MultimodalEmotionlinesDataset
Dataset Card for "emotion_recognition_multimodal_emotionlines_dataset"
More Information needed
accelerometer_raw_emotion_recognition_smartphone_sensorsEmotion Recognition ML on the Basis of Smartphone Accelerometer Sensors ✨
Hi,
I tried to find a correlation between emotional states (or vibes, as they are referred to in the deployed mobile applications) and the way the user moves the phone while typing (for example, a message). I wanted to represent these emotions then in color (an accompanying color bubble after each message on WhatsApp, for example).
I collected data from myself in different emotional states, ending up with around an hour… See the full description on the dataset page: https://huggingface.co/datasets/wolfeiq/accelerometer_raw_emotion_recognition_smartphone_sensors.gn-emotion-recognition
Text-based afective computing
We collected a dataset of tweets primarily written in Guarani (and Jopara, a code-switching language that combines Guarani and Spanish) and annotated them for three widely-used dimensions in sentiment analysis:
emotion recognition (this repo, https://huggingface.co/datasets/mmaguero/gn-emotion-recognition),
humor detection (https://huggingface.co/datasets/mmaguero/gn-humor-detection), and
offensive language identification… See the full description on the dataset page: https://huggingface.co/datasets/mmaguero/gn-emotion-recognition.CASIA_speech_emotion_recognitionSpeech_Emotion_RecognitionPromptTTS_Emotion_Recognition
Dataset Card for "PromptTTS_Emotion_Recognition"
More Information needed
CASIA_speech_emotion_recognition_preloadFace_Emotion_Insight_Recognition_Dataset
Multimodal Emotion & Physiological Analysis
Project Overview
This project explores the relationship between physiological signals and facial micro expressions to improve emotion recognition in therapeutic settings. This research assists in determining which facial and physiological signals should be prioritized to detect clinical "misalignment" or hidden distress during therapy sessions.
Dataset Selection & Description
Source: The dataset is sourced from Kaggle (Face Emotion & Physiological… See the full description on the dataset page: https://huggingface.co/datasets/ofekponzo/Face_Emotion_Insight_Recognition_Dataset.Emotion_Recognition_4_llama2
Dataset Card for "Emotion_Recognition_4_llama2"
More Information needed
Emotion_Recognition_4_llama2_chat_oversampled
Dataset Card for "Emotion_Recognition_4_llama2_chat_oversampled"
More Information needed
gender_emotion_recognitionEmoji-Grounded_Speech_Emotion_Recognition
Dataset Card for "Emoji-Grounded_Speech_Emotion_Recognition"
More Information needed
emotion-recognition-datasetEmotionRecognition_MultimodalEmotionlinesDataset_TTSEmotion_Recognition_4_llama2_v3
Dataset Card for "Emotion_Recognition_4_llama2_v3"
More Information needed
SLT-Task3-Post-ASR-Emotion-Recognition
Dataset Name: Text-Only ASR transcripts of IEMOCAP for ASR error correction and emotion recognition
This is an n-best hypotheses augmented dataset of IEMOCAP, which is used for ASR error correction and emotion recognition.
Please see IEMOCAP first to access the raw audio dataset.
Description
This dataset consists of ASR transcripts of 11 speech models, following the turns of the conversation in IEMOCAP, with corresponding speaker ID and utterance ID.
To acquire this… See the full description on the dataset page: https://huggingface.co/datasets/GenSEC-LLM/SLT-Task3-Post-ASR-Emotion-Recognition.EmotionRecognition_MultimodalEmotionlinesDataset_TTSEmotion_Recognition_4_llama2_v2
Dataset Card for "Emotion_Recognition_4_llama2_v2"
More Information needed
