CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01google-research-datasets /go_emotions Dataset Card for GoEmotions Dataset Summary The GoEmotions dataset contains 58k carefully curated Reddit comments labeled for 27 emotion categories or Neutral. The raw data is included as well as the smaller, simplified version of the dataset with predefined train/val/test splits. Supported Tasks and Leaderboards This dataset is intended for multi-class, multi-label emotion classification. Languages The data is in English. Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/google-research-datasets/go_emotions.tabulartext-classification100K<n<1M267 likes14k downloads3y agoHugging Face02ChristophSchuhmann /emotionsimage100K<n<1M2 likes3.9k downloads3y agoHugging Face03brighter-dataset /BRIGHTER-emotion-categories BRIGHTER Emotion Categories Dataset This dataset contains the emotion categories data from the BRIGHTER paper: BRIdging the Gap in Human-Annotated Textual Emotion Recognition Datasets for 28 Languages. Dataset Description The BRIGHTER Emotion Categories dataset is a comprehensive multi-language, multi-label emotion classification dataset with separate configurations for each language. It represents one of the largest human-annotated emotion datasets across multiple… See the full description on the dataset page: https://huggingface.co/datasets/brighter-dataset/BRIGHTER-emotion-categories.tabular100K<n<1M18 likes1.8k downloads10mo agoHugging Face04pollen-robotics /microduck-emotions Microduck Emotions A collection of emotions for the Microduck robot. Each one is a motion and a sound designed together, beat by beat, with the beak opening on the sound, rendered in the physics simulation and validated on the real robot. Every emotion is three files: the motion (emotions/<name>.json, keyframes at 30 fps: head and body offsets played on top of whichever trained policy is active, plus the policy hand-overs, such as the sit that devastated and play dead start)… See the full description on the dataset page: https://huggingface.co/datasets/pollen-robotics/microduck-emotions.audioroboticsn<1K6 likes979 downloads19d agoHugging Face05laion /laion-emotional-trajectory-t80 LAION Emotional-Trajectory Speech — tier T≥0.80 319,765 crossfaded speech trajectories · 4,482 audio-hours · 1,598,825 source clips A trajectory is a short sequence of 5 consecutive utterances by one speaker whose measured emotion or voice character moves monotonically from one end of the corpus distribution to the other. The clips are joined into one continuous audio file with equal-power crossfades, the joined audio is re-tokenized with MOSS-Audio- Tokenizer-v2, and every… See the full description on the dataset page: https://huggingface.co/datasets/laion/laion-emotional-trajectory-t80.tabulartext-to-speech100K<n<1M0 likes607 downloads17d agoHugging Face06SetFit /go_emotions GoEmotions This dataset is a port of the official go_emotions dataset on the Hub. It only contains the simplified subset as these are the only fields we need for text classification. tabular10K<n<100K13 likes591 downloads4y agoHugging Face07ukr-detect /ukr-emotions-binary EmoBench-UA: Emotions Detection Dataset in Ukrainian Texts EmoBench-UA: the first of its kind emotions detection dataset in Ukrainian texts. This dataset covers the detection of basic emotions: Joy, Anger, Fear, Disgust, Surprise, Sadness, or None. Any text can contain any amount of emotion -- only one, several, or none at all. The texts with None emotions are the ones where the labels per emotions classes are 0. Binary: specifically this dataset contains binary labels… See the full description on the dataset page: https://huggingface.co/datasets/ukr-detect/ukr-emotions-binary.imagetext-classification1K<n<10K0 likes293 downloads2mo agoHugging Face08seara /ru_go_emotions Description This dataset is a translation of the Google GoEmotions emotion classification dataset. All features remain unchanged, except for the addition of a new ru_text column containing the translated text in Russian. For the translation process, I used the Deep translator with the Google engine. You can find all the details about translation, raw .csv files and other stuff in this Github repository. For more information also check the official original dataset card.… See the full description on the dataset page: https://huggingface.co/datasets/seara/ru_go_emotions.tabulartext-classification100K<n<1M19 likes244 downloads3y agoHugging Face09brighter-dataset /BRIGHTER-emotion-intensities BRIGHTER Emotion Intensities Dataset This dataset contains the emotion intensities data from the BRIGHTER paper: BRIdging the Gap in Human-Annotated Textual Emotion Recognition Datasets for 28 Languages. Dataset Description The BRIGHTER Emotion Intensities dataset is a comprehensive multi-language emotion intensity dataset with separate configurations for each language. It represents one of the largest human-annotated emotion datasets across multiple languages, providing… See the full description on the dataset page: https://huggingface.co/datasets/brighter-dataset/BRIGHTER-emotion-intensities.tabular10K<n<100K5 likes224 downloads11mo agoHugging Face10m-a-p /EMOaudion<1K1 likes199 downloads1y agoHugging Face11BAAI /IndustryInstruction_Literature-Emotions IndustryInstruction: Literature & Emotions This repository contains the IndustryInstruction: Literature & Emotions domain subset of BAAI/IndustryInstruction. Refer to the parent dataset card for data construction, intended use, limitations, and licensing details. Citation If you use this dataset in your work, please cite IndustryInstruction: @misc{shi2024industryinstruction, title = {IndustryInstruction}, author = {Xiaofeng Shi and Lulu Zhao and Hua… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryInstruction_Literature-Emotions.tabularquestion-answering100K<n<1M2 likes189 downloads1mo agoHugging Face12mahalisyarifuddin /emotweetid-ekman7 EmoTweetID under Ekman's seven universal emotions 2,243 Indonesian tweets, one label each from Ekman's universal set - anger, contempt, disgust, enjoyment, fear, sadness, surprise. 475 of them are the pool EmoTweetID tagged anger, and that pool is the only place this dataset makes a decision of its own: laya reads each of those tweets, in Indonesian, against Ekman's own definitions of the two emotions, and splits them into anger (289) and contempt (186) - contempt is 39.2% of… See the full description on the dataset page: https://huggingface.co/datasets/mahalisyarifuddin/emotweetid-ekman7.tabulartext-classification1K<n<10K0 likes178 downloads6d agoHugging Face13llama-lang-adapt /EmotionAnalysisFinal Dataset Card for EmotionAnalysisFinal EmotionAnalysisFinal is the official dataset for SemEval-2025 Task 11, Track C: Cross-lingual Emotion Detection in Social Media Text. This dataset comprises multilingual social media posts annotated for six basic emotions: anger, disgust, fear, joy, sadness, and surprise. The annotation schema is multi-label. Each language-specific configuration (subset) contains validation and test splits. Split Original Source Notes validation dev… See the full description on the dataset page: https://huggingface.co/datasets/llama-lang-adapt/EmotionAnalysisFinal.tabulartext-classification10K<n<100K0 likes166 downloads1y agoHugging Face14stepp1 /tweet_emotion_intensity Tweet Emotion Intensity Dataset Papers: Emotion Intensities in Tweets. Saif M. Mohammad and Felipe Bravo-Marquez. In Proceedings of the sixth joint conference on lexical and computational semantics (*Sem), August 2017, Vancouver, Canada. WASSA-2017 Shared Task on Emotion Intensity. Saif M. Mohammad and Felipe Bravo-Marquez. In Proceedings of the EMNLP 2017 Workshop on Computational Approaches to Subjectivity, Sentiment, and Social Media (WASSA), September 2017… See the full description on the dataset page: https://huggingface.co/datasets/stepp1/tweet_emotion_intensity.tabular1K<n<10K9 likes161 downloads4y agoHugging Face15roboticshack /team9-lepuppet_emotionstabular10K<n<100K0 likes153 downloads1y agoHugging Face16HSP-IIT /eval_groot_emo_demoThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "ergocub", "total_episodes": 0, "total_frames": 0, "total_tasks": 0, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 10, "splits": {}, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/HSP-IIT/eval_groot_emo_demo.tabularrobotics10K<n<100K0 likes151 downloads3d agoHugging Face17humanlong /emotion-negotiation-benchmarks Emotion-Aware LLM Negotiation Benchmarks Four high-stakes, edge-deployable negotiation benchmarks — the official evaluation suite for our research program on emotion-aware LLM agents. Each benchmark targets a distinct domain where (a) LLM-vs-LLM negotiation has real-world consequences, and (b) on-device deployment of small language models matters for privacy and latency. The benchmarks were originally introduced with EmoMAS (ACL 2026 Main, top 9% of 12,148 submissions) and are… See the full description on the dataset page: https://huggingface.co/datasets/humanlong/emotion-negotiation-benchmarks.tabulartext-generationn<1K0 likes141 downloads4mo agoHugging Face18cagataydev /emotiv-ecot emotiv-ecot Embodied Chain-of-Thought episodes where the body is a human cortex: one person in an EMOTIV EPOC X talking to an agent that reads a one-line brain summary before every reply. Each turn becomes a LeRobot v3.0 episode (Zawalski et al. 2024 with the robot body swapped for a head): the brain is observation and reward (Δstress, Δengagement across the reply), the agent's speech is the action, the reasoning is the per-frame TASK | AMBIENT | PLAN | TOOL | ACT | REWARD… See the full description on the dataset page: https://huggingface.co/datasets/cagataydev/emotiv-ecot.tabularroboticsn<1K0 likes120 downloads24d agoHugging Face19llm-council /emotional_application Data explorer and full leaderboard https://huggingface.co/spaces/llm-council/emotional-intelligence-arena The LMC-EA dataset This dataset was developed to demonstrate how to benchmark foundation models on highly subjective tasks such as those in the domain of emotional intelligence by the collective consensus of a council of LLMs. There are 4 subsets of the LMC-EA dataset: test_set_formulation: Synthetic expansions of the EmoBench EA dataset, generated by 20… See the full description on the dataset page: https://huggingface.co/datasets/llm-council/emotional_application.tabular10K<n<100K6 likes106 downloads2y agoHugging Face20knoveleng /emotion-datasets Emotion datasets Synthetic emotion text re-generated from the data pipelines of Emotion concepts and their function in a LLM (paper), for interpretability and steering research. This is a re-generation with a different model, not the paper authors' data; prompts, the 171-emotion word list and the 100 story topics come from the paper's appendix. Total: 4,061 rows across 4 configs. config rows what it is stories 2,718 one story per row, one target emotion each (12… See the full description on the dataset page: https://huggingface.co/datasets/knoveleng/emotion-datasets.tabulartext-generation1K<n<10K0 likes95 downloads6d agoHugging Face21myned-ai /audio2face-emotion-arkit-teacher audio2face-emotion-arkit-teacher Nyx avatar (Gaussian-splat head, ARKit-52 blendshape rig) driven by a surprise clip's blendshape labels derived from this dataset. 14,082 emotional-speech clips, each annotated with two parallel 52-channel ARKit blendshape sequences (NVIDIA Audio2Face-3D-v2.3.1-James and LAM_Audio2Expression) plus a 26-dimensional NVIDIA Audio2Emotion conditioning vector. Reference-only dataset — the original audio is not shipped. Each row contains a… See the full description on the dataset page: https://huggingface.co/datasets/myned-ai/audio2face-emotion-arkit-teacher.tabularaudio-classification10K<n<100K3 likes92 downloads4mo agoHugging Face22ukr-detect /ukr-emotions-intensity EmoBench-UA: Emotions Detection Dataset in Ukrainian Texts EmoBench-UA: the first of its kind emotions detection dataset in Ukrainian texts. This dataset covers the detection of basic emotions: Joy, Anger, Fear, Disgust, Surprise, Sadness, or None. Any text can contain any amount of emotion -- only one, several, or none at all. The texts with None emotions are the ones where the labels per emotions classes are 0. Intensity: specifically this dataset contains intensity labels… See the full description on the dataset page: https://huggingface.co/datasets/ukr-detect/ukr-emotions-intensity.imagetext-classification1K<n<10K0 likes89 downloads2mo agoHugging Face23FatimahEmadEldin /Moroccan-Arabic-Multimodal-Emotion-Recognition MDER-MA — Moroccan Arabic Multimodal Emotion Recognition (TTS-aligned repackaging) A repackaging of the MDER-MA dataset that pairs every audio clip with its Arabic (Moroccan dialect / Darija) transcript and ships speaker-disjoint train/validation/test splits. Original dataset: Ouali, S. & El Garouani, S. (2025). MDER-MA: A multimodal dataset for emotion recognition in low-resource Moroccan Arabic language. Data in Brief. DOI: 10.1016/j.dib.2025.112005. Mendeley:… See the full description on the dataset page: https://huggingface.co/datasets/FatimahEmadEldin/Moroccan-Arabic-Multimodal-Emotion-Recognition.audiotext-to-speech1K<n<10K1 likes86 downloads5mo agoHugging Face24laion /emolia-voicenet-gemini-annotations Emolia VoiceNet Gemini Annotations 468,180 dimension-level annotations over 236,613 Emolia speech clips, each scored 0-6 (0-2 for the content-safety dimension) on one of 57 perceptual voice / speech dimensions - arousal, valence, brightness, resonance placement, speaking styles, genuineness, recording quality, and more - by Gemini 3.5 Flash (non-thinking, temperature 0). This repository ships the annotations, audio provenance, per-dimension statistics, and the full scoring… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-voicenet-gemini-annotations.tabularaudio-classification100K<n<1M0 likes78 downloads2mo agoHugging Face25WhaleDolphin /MIKU-EmoBench MIKU-PAL/MIKU-EmoBench: An Automatic Multi-Modal Method for Audio Paralinguistic and Affect Labeling This is the official repository for the MIKU-EmoBench dataset annotations. MIKU-EmoBench is a novel, large-scale dataset specifically designed for audio paralinguistic and affect labeling, addressing critical limitations of existing emotional datasets in terms of scale and granularity. Developed using our MIKU-PAL pipeline, MIKU-EmoBench rapidly collected about 160 hours of… See the full description on the dataset page: https://huggingface.co/datasets/WhaleDolphin/MIKU-EmoBench.tabular10K<n<100K2 likes71 downloads1y agoHugging Face26llm-lab /Emo3D Emo3D: Metric and Benchmarking Dataset for 3D Facial Expression Generation from Emotion Description Citation @inproceedings{dehghani-etal-2025-emo3d, title = "{E}mo3{D}: Metric and Benchmarking Dataset for 3{D} Facial Expression Generation from Emotion Description", author = "Dehghani, Mahshid and Shafiee, Amirahmad and Shafiei, Ali and Fallah, Neda and Alizadeh, Farahmand and Gholinejad, Mohammad Mehdi and Behroozi, Hamid… See the full description on the dataset page: https://huggingface.co/datasets/llm-lab/Emo3D.image10K<n<100K1 likes68 downloads1y agoHugging Face27slone /e-mordovia-articles-2024 "e-mordovia-articles-2024": a parallel news dataset for Russian, Erzya and Moksha This is a semi-aligned dataset of Russian, Erzya and Moksha news articles, crawled from https://www.e-mordovia.ru. Dataset Description Dataset Summary This is a dataset of news articles collected from https://www.e-mordovia.ru, the official portal of the state authorities of the Republic of Mordovia. The articles have been paired by the following algorithm: Calculate similarities… See the full description on the dataset page: https://huggingface.co/datasets/slone/e-mordovia-articles-2024.tabulartranslation100K<n<1M3 likes67 downloads1y agoHugging Face28ss-emoz /items_raw_fulltabular100K<n<1M0 likes67 downloads7mo agoHugging Face29PSewmuthu /Emotion_Video_Facial_Landmarks Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset Dataset Description This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files. License: The extracted feature data in this CSV file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/PSewmuthu/Emotion_Video_Facial_Landmarks.tabularimage-classification100K<n<1M0 likes64 downloads11mo agoHugging Face30Sammaiah /MELD-processed-v5-emotion2vec MELD Processed Multi-Modal Emotion Recognition Dataset Processed dataset containing Prosody, Whisper acoustic encodings, DistilBERT text hidden states, and Ekman emotion labels. tabular10K<n<100K0 likes61 downloads7d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.