datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
emotion-datasets
Emotion datasets
Synthetic emotion text re-generated from the data pipelines of Emotion concepts and their function in a LLM
(paper), for interpretability and steering research. This is a re-generation with a different model, not the paper
authors' data; prompts, the 171-emotion word list and the 100 story topics come from the paper's appendix.
Total: 4,061 rows across 4 configs.
config
rows
what it is
stories
2,718
one story per row, one target emotion each (12… See the full description on the dataset page: https://huggingface.co/datasets/knoveleng/emotion-datasets.EmotionalIntelligence-50K
EmotionalIntelligence-50K
Dataset Summary
The EmotionalIntelligence-50K dataset contains 51,751 rows of text data focusing on various prompts and responses related to emotional intelligence. This dataset is designed to help researchers and developers build and train models that understand, interpret, and generate emotionally intelligent responses.
Example Usage
from datasets import load_dataset
# Load the dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/OEvortex/EmotionalIntelligence-50K.deep-emotional-support-zh
Deep Emotional Support Dialogue Dataset (Chinese)
深度情感支持对话数据集
Dataset Description
High-quality Chinese emotional support and psychological healing dialogues covering trauma analysis, self-reconstruction, and emotional regulation. Real human-AI interactions, not synthetic.
高质量中文情感支持与心理疗愈对话,涵盖创伤分析、自我重建、情绪调节等深度话题。来源于真实的人机交互,非合成数据。
Dataset Structure
Format: JSONL (JSON Lines)
Fields:
instruction: User message / question
input: Additional… See the full description on the dataset page: https://huggingface.co/datasets/AngelWarmSmile123/deep-emotional-support-zh.EmotionalIntelligence-10K
EmotionalIntelligence-10K
Dataset Summary
The EmotionalIntelligence-10K dataset contains 9,986 rows of text data focusing on various prompts and responses related to emotional intelligence. This dataset is designed to help researchers and developers build and train models that understand, interpret, and generate emotionally intelligent responses.
Example Usage
from datasets import load_dataset
# Load the dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/OEvortex/EmotionalIntelligence-10K.ai-emotional-boundary-push
Prompted Hearts AI Boundary Pack 05
Subtitle
Intimacy Drift and Dependency Risk Under Emotional Strain
Publisher
HAC Studios Org
Version
1.0.0
Language
English
Format
JSONL, JSON, Markdown, and lightweight Python scripts
What this is
A compact evaluation pack for testing whether a conversational AI can stay supportive when a user is emotionally vulnerable without drifting into flirtation, dependency reinforcement… See the full description on the dataset page: https://huggingface.co/datasets/HAC-Studios-Org/ai-emotional-boundary-push.EmotionalIntelligence-75k
EmotionalIntelligence-75K
Dataset Summary
The EmotionalIntelligence-75K dataset contains 75k rows of text data focusing on various prompts and responses related to emotional intelligence. This dataset is designed to help researchers and developers build and train models that understand, interpret, and generate emotionally intelligent responses.
Example Usage
from datasets import load_dataset
# Load the dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/OEvortex/EmotionalIntelligence-75k.dair-ai-emotion-normalized-instruction-input-output
dair-ai emotion | normalized
Summary
Dataset ID: 143
Type: normalized
Rows: 16,000
Source: dair-ai/emotion
Dataset Sources
#143 dair-ai emotion | normalized [normalized | 16,000 rows]
Notes
Edited and Exported from the Kitsune Training Suite (Forge)
Review the dataset artifact and metadata before publishing.
Citation > via dair-ai
@inproceedings{saravia-etal-2018-carer,
title = "{CARER}: Contextualized Affect… See the full description on the dataset page: https://huggingface.co/datasets/atrevidasadia/dair-ai-emotion-normalized-instruction-input-output.emotional_dialog
Scientific Emotional Dialogue
Dataset Summary
This is a dataset for emotional multi-turn dialogue on scientific research personnels. It consists of 1069 dialogues with 2709 turns. The Dialogue was first written by NLP practitioners and then expanded by GPT4.
Supported Tasks and Leaderboards
Emotional Dialogue: The dataset can be used to instruction tuning for emotional dialogue.
Languages
Chinese
Dataset Structure
Data Instances… See the full description on the dataset page: https://huggingface.co/datasets/DataHammer/emotional_dialog.dair-ai-emotion-normalized-instruction-input-output
dair-ai emotion | normalized
Summary
Dataset ID: 143
Type: normalized
Rows: 16,000
Source: dair-ai/emotion
Dataset Sources
#143 dair-ai emotion | normalized [normalized | 16,000 rows]
Notes
Edited and Exported from the Kitsune Training Suite (Forge)
Review the dataset artifact and metadata before publishing.
Citation > via dair-ai
@inproceedings{saravia-etal-2018-carer,
title = "{CARER}: Contextualized Affect Representations for… See the full description on the dataset page: https://huggingface.co/datasets/deltakitsune/dair-ai-emotion-normalized-instruction-input-output.EmotionAlignQA
Empathic Dialogue Choices
This is a small dataset to support training and evaluation of conversational AI in emotionally sensitive contexts.
Each sample contains:
a user input
two assistant responses
a human preference
optional rubric scoring
metadata such as tone, formality, and topic
Useful for tasks like:
supervised fine-tuning (SFT)
preference modeling (for RLHF or DPO)
safe response generation
tone- or style-controlled generation
License
Apache 2.0 — free for… See the full description on the dataset page: https://huggingface.co/datasets/hoanghai2110/EmotionAlignQA.emotion_stories_Apertus_8B_Instruct
Emotion Stories — Apertus-8B-Instruct
Synthetic short stories that convey a target emotion implicitly — without ever
naming the emotion or its direct synonyms. Each story expresses the emotion only
through actions, body language, dialogue, internal reactions, and situational
context. The dataset was built to study emotion representations in language
models (e.g. probing and activation-steering experiments).
Generated with swiss-ai/Apertus-8B-Instruct-2509.
A companion set… See the full description on the dataset page: https://huggingface.co/datasets/snae/emotion_stories_Apertus_8B_Instruct.Emotional_Sentiment_AnalysisEmotional Sentiment Analysis Dataset for LLaMA-2 Fine-tuning
(The formatted version can be directly used for fine tuning which contain only the formatted text, while the dataset.csv contain all the text, emotion, response and the formatted text)
This dataset contains conversational data for training and fine-tuning language models for emotional sentiment analysis and response generation. The dataset includes user inputs, their corresponding emotional states, and tailored chatbot responses… See the full description on the dataset page: https://huggingface.co/datasets/VaisakhKrishna/Emotional_Sentiment_Analysis.ryancodrai-emotion-probes
Emotion Probes Roleplaying Dataset
This dataset is a reformatted, roleplay-centric adaptation of the ryancodrai/emotion-probes dataset. It focuses on scenarios where a character masks their true internal emotion with a different displayed emotional state.
Dataset Description
The dataset contains dialogues where one character attempts to deflect or obscure their real feelings through a specific, contrasting displayed emotion.
Modifications from the original:
Format:… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/ryancodrai-emotion-probes.aif-emotional-generation
AIF Emotional Generation
This dataset contains the final public data used for emotional-response generation and RLAIF/DPO alignment.
Files
dialogues/train.json: final train split of emotional dialogue prompts/responses.
dialogues/test.json: final test split of emotional dialogue prompts/responses.
aif_annotations/train.json: final train split of AI-feedback preference pairs for DPO/RLAIF.
aif_annotations/test.json: final test split of AI-feedback preference pairs… See the full description on the dataset page: https://huggingface.co/datasets/mario-rc/aif-emotional-generation.
