datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Robot-EQ
RobotEQ-Data
Official dataset release for RobotEQ.
Evaluation & Scripts
For inference scripts, evaluation scripts, and data production tooling, see the RobotEQ code repository.
Dataset Statistics
Item
Count
Behavior judgment scenarios (synthetic)
1,812
Behavior judgment scenarios (real POV)
223
Behavior judgment scenarios (total)
2,035
Behavior judgment behavior annotations
3,171
Spatial grounding questions
825… See the full description on the dataset page: https://huggingface.co/datasets/Tongji-Emotion/Robot-EQ.IndustryInstruction_Literature-Emotions
IndustryInstruction: Literature & Emotions
This repository contains the IndustryInstruction: Literature & Emotions domain subset of BAAI/IndustryInstruction.
Refer to the parent dataset card for data construction, intended use, limitations,
and licensing details.
Citation
If you use this dataset in your work, please cite IndustryInstruction:
@misc{shi2024industryinstruction,
title = {IndustryInstruction},
author = {Xiaofeng Shi and Lulu Zhao and Hua… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryInstruction_Literature-Emotions.EmoBench
EmoBench
This is the official repository for our ACL 2024 paper "EmoBench: Evaluating the Emotional Intelligence of Large Language Models"
Overview
EmoBench is a comprehensive and challenging benchmark designed to evaluate the Emotional Intelligence (EI) of Large Language Models (LLMs). Unlike traditional datasets, EmoBench focuses not only on emotion recognition but also on advanced EI capabilities such as emotional reasoning and application.
The dataset includes 400… See the full description on the dataset page: https://huggingface.co/datasets/SahandSab/EmoBench.Psych8kThe data used for this project comes from ~260 real conversations in counseling recordings (in English). The transcripts of these recordings were used as the primary source for building the training and testing datasets. These conversations cover a variety of topics, including emotion, family, relationship, career development, academic stress, etc
Recent News
[2024-04-07] After obtaining access permission, please do not disseminate data at will!
telelogs
TeleLogs Dataset (Processed MCQ Format)
This dataset has been extracted from the original netop/TeleLogs dataset and processed into multiple-choice question (MCQ) format for easier evaluation.
Dataset Description
TeleLogs is a telecommunications log analysis benchmark where models must identify the root cause of network issues from 5G wireless network drive-test data and engineering parameters.
Processed Format
This version has been restructured for MCQ… See the full description on the dataset page: https://huggingface.co/datasets/emolero/telelogs.EmoSupportBench
EmoSupportBench
EmoSupportBench is a comprehensive dataset and benchmark for evaluating emotional support capabilities of large language models (LLMs). It provides a systematic framework to assess how well AI systems can provide empathetic, helpful, and psychologically-grounded support to users seeking emotional assistance.
🎯 Key Features
200-question bilingual evaluation set (English & Chinese) covering 8 major emotional support scenarios
Hierarchical scenario… See the full description on the dataset page: https://huggingface.co/datasets/YueyangWang/EmoSupportBench.telelogs_markdown
TeleLogs Dataset (Processed MCQ Format)
This dataset has been extracted from the original netop/TeleLogs dataset and processed into multiple-choice question (MCQ) format for easier evaluation.
Dataset Description
TeleLogs is a telecommunications log analysis benchmark where models must identify the root cause of network issues from 5G wireless network drive-test data and engineering parameters.
Processed Format
This version has been restructured for MCQ… See the full description on the dataset page: https://huggingface.co/datasets/emolero/telelogs_markdown.BizFinBench
BizFinBench: A Business-Driven Real-World Financial Benchmark for Evaluating LLMs
📖Paper |🐙Github|🤗Huggingface
Large language models excel in general tasks, yet assessing their reliability in logic‑heavy, precision‑critical domains like finance, law, and healthcare remains challenging. To address this, we introduce BizFinBench, the first benchmark specifically designed to evaluate LLMs in real-world financial applications. BizFinBench comprises over 100,000+ bilingual (English &… See the full description on the dataset page: https://huggingface.co/datasets/emoxaj/BizFinBench.EmoPropManalpaca-QA-conciousness-emotions
Dataset Card for Dataset Name
This is a concious thoughts and emotions thought dataset from Gemini 1.5 flash synthethic data.
Dataset Details
conciousness
emotions
thoughts from AI
Uses
Fine tune with Alpaca format.
alpaca_prompt = """Below is an instruction that describes a task, paired with an input that provides further context. Write a response that appropriately completes the request.
### Instruction:
{}
### Input:
{}
### Response:
{}"""
EOS_TOKEN… See the full description on the dataset page: https://huggingface.co/datasets/EpistemeAI/alpaca-QA-conciousness-emotions.Ques-Ans-with-Emotioncreated and shared by: 0xarchit
Q-A-M-khmer-Emotion
Khmer Emotion QA Dataset
Welcome to the SeyhaLite collection. This dataset has been meticulously curated and cleaned to support the development of high-quality Khmer Language Models (LLMs) and Question-Answering systems.
Project Vision
I hope this dataset helps your project succeed. Whether you are building a chatbot, an assistant, or conducting research, this data is designed to provide clear and accurate information about emotional expressions and sentiment in the Khmer… See the full description on the dataset page: https://huggingface.co/datasets/SeyhaLite/Q-A-M-khmer-Emotion.
