datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Robot-EQ
RobotEQ-Data
Official dataset release for RobotEQ.
Evaluation & Scripts
For inference scripts, evaluation scripts, and data production tooling, see the RobotEQ code repository.
Dataset Statistics
Item
Count
Behavior judgment scenarios (synthetic)
1,812
Behavior judgment scenarios (real POV)
223
Behavior judgment scenarios (total)
2,035
Behavior judgment behavior annotations
3,171
Spatial grounding questions
825… See the full description on the dataset page: https://huggingface.co/datasets/Tongji-Emotion/Robot-EQ.IndustryInstruction_Literature-Emotions
IndustryInstruction: Literature & Emotions
This repository contains the IndustryInstruction: Literature & Emotions domain subset of BAAI/IndustryInstruction.
Refer to the parent dataset card for data construction, intended use, limitations,
and licensing details.
Citation
If you use this dataset in your work, please cite IndustryInstruction:
@misc{shi2024industryinstruction,
title = {IndustryInstruction},
author = {Xiaofeng Shi and Lulu Zhao and Hua… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryInstruction_Literature-Emotions.Emotions-Annotated-Customer-Care-QA-Dataset-Romanized-and-Devanagari
Dataset Card for Dataset Name
यो देवनागरी नेपाली भाषाको डेटासेट विशेषगरी च्याटबोट प्रणालीहरू बनाउनको लागि डिजाइन गरिएको हो। यसमा विभिन्न श्रेणीहरूको डेटासेटहरू समावेश गरिएको छ, जसलाई JSON मा ढाँचा बनाईएको छ, जसले नेपाली वार्तालाप एआई अनुप्रयोगहरूको लागि भाषा मोडेलहरूलाई तालिम र फाइन-ट्यून गर्नको लागि व्यापक स्रोत प्रदान गर्दछ।
Dataset Prepared by:
Manoj Kumar Baniya
Aakash Kumar Thakur
Manish Kathet
Kshitiz Gajurel
Dataset Details
Dataset Description… See the full description on the dataset page: https://huggingface.co/datasets/kshitizgajurel/Emotions-Annotated-Customer-Care-QA-Dataset-Romanized-and-Devanagari.lora-emotional-alignment-sample
BrightRun BrightRun Emotional Alignment Dataset — Sample Preview
🎯 Train Your LLM to Handle Emotionally Complex Conversations
This is a 12-conversation sample. The full dataset contains 242 conversations and 1,567 training pairs.
⚠️ This is a Sample — Not the Full Dataset
You're looking at 12 sample conversations designed to help you evaluate data quality before downloading the complete dataset.
What You Get Here
What You Get at brighthub.ai… See the full description on the dataset page: https://huggingface.co/datasets/BrightHubAI/lora-emotional-alignment-sample.emotionlessfull
EmotionLessFull
Experiment on how smaller, dumber LLMs can get its safeguards bypassed into NSFW-like conversations.
alpaca-QA-conciousness-emotions
Dataset Card for Dataset Name
This is a concious thoughts and emotions thought dataset from Gemini 1.5 flash synthethic data.
Dataset Details
conciousness
emotions
thoughts from AI
Uses
Fine tune with Alpaca format.
alpaca_prompt = """Below is an instruction that describes a task, paired with an input that provides further context. Write a response that appropriately completes the request.
### Instruction:
{}
### Input:
{}
### Response:
{}"""
EOS_TOKEN… See the full description on the dataset page: https://huggingface.co/datasets/EpistemeAI/alpaca-QA-conciousness-emotions.Ques-Ans-with-Emotioncreated and shared by: 0xarchit
Q-A-M-khmer-Emotion
Khmer Emotion QA Dataset
Welcome to the SeyhaLite collection. This dataset has been meticulously curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) and Question-Answering systems.
Project Vision
I hope this dataset helps your project succeed. Whether you are building a chatbot, an assistant, or conducting research, this data is designed to provide clear and accurate information about emotional expressions and sentiment in the Khmer… See the full description on the dataset page: https://huggingface.co/datasets/SeyhaLite/Q-A-M-khmer-Emotion.
