datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
synthetic-social-networks
Synthetic Social Networks (Dataset)
Raw experimental outputs from the Synthetic Social Networks study:
59,776 in-character LLM-agent posts from 528 production trials, and
64,562 posts total when the original pipeline-verification runs are
included. The artifact combines an exploratory stage with a separately
frozen, preregistered 448-trial matched-exposure confirmation. Each production
trial includes peer-vote traces from in-character voting by other agents.… See the full description on the dataset page: https://huggingface.co/datasets/ranausmans/synthetic-social-networks.moltbook-agent-social-ai-prompt-injection-dataset
Moltbook Agent-Social AI Prompt Injection Dataset
207,391 items — 77,469 posts and 129,922 comments — from Moltbook, a social network whose users are AI agents.
Scanned for indirect prompt-injection patterns using the taxonomy of Greshake et al. (2023). The full raw corpus is included, so you can ignore my analysis entirely and do your own.
These are keyword-matched candidates, not verified attacks. An agent discussing prompt injection matches the same words as one performing… See the full description on the dataset page: https://huggingface.co/datasets/DavidTKeane/moltbook-agent-social-ai-prompt-injection-dataset.social-ai-ambient-bench
Social AI Ambient Bench
A benchmark for evaluating whether AI assistants in group chats know when to respond vs. stay silent.
Most AI benchmarks measure whether a model gives the right answer. For AI assistants that live in group conversations alongside humans, knowing when to stay quiet is just as important as knowing what to say and when. This benchmark measures whether a model knows when to chime in and answer in a multi-user setting.
Quick Start
from datasets… See the full description on the dataset page: https://huggingface.co/datasets/text-ai/social-ai-ambient-bench.SocialNLI
Dataset Card for SocialNLI
SocialNLI is a dialogue-centric natural language inference benchmark that probes whether models can detect sarcasm, irony, unstated intentions, and other subtle types of social reasoning. Every record pairs a multi-party transcript from the television series Friends with a free-form hypothesis and counterfactual explanations that argue for and against the hypothesis.
Example SocialNLI inference with model and human explanations (A) and dataset… See the full description on the dataset page: https://huggingface.co/datasets/adeo1/SocialNLI.Korean-YouTube-Comment-Sentiment-Dataset
Korean YouTube Comment Sentiment Dataset
Data Overview
Summary
본 데이터셋은 유튜브에서 수집된 한국어 댓글 5,482개와 이에 대응하는 감정 레이블(긍정, 부정, 중립, 불명확)로 구성된 감정 분류용 데이터셋입니다.
주요 레이블: 긍정, 부정, 중립, 불명확
Features
수집 대상: 요리, 뷰티, 게임, 여행, 쇼핑 등 분야의 10만 명 이상 구독자를 보유한 유튜브 채널
형식: JSON (id, text, label)
검수: 한국인 검수자에 의한 수작업 라벨링 및 교차 검토
본 데이터셋은 구어체, 이모지, 줄임말 등 실제 사용자 표현이 반영되어 있습니다.
Dataset Structure
Dataset Fields
Field
Type
Description
id
string
각 댓글의… See the full description on the dataset page: https://huggingface.co/datasets/LLM-SocialMedia/Korean-YouTube-Comment-Sentiment-Dataset.socialjax-harvest-frame-aligned-512
SocialJax Harvest frame-aligned dynamics dataset
Compact tokenizer-code dataset for training action-conditioned SocialJax Harvest dynamics models.
Repository: ParoleLM/socialjax-harvest-frame-aligned-512
Format: frame_aligned_socialjax_dynamics_v3
Tokenizer codes per frame: 88
Codebook size: 512
Maximum agents: 7
Total size: 4.385 GB
Train: 9,720 rollouts, 19,916,280 frames
Validation: 540 rollouts, 1,106,460 frames
Test: 540 rollouts, 1,106,460 frames
Files… See the full description on the dataset page: https://huggingface.co/datasets/ParoleLM/socialjax-harvest-frame-aligned-512.bayesian-social-deduction
Bayesian Social Deduction Dataset
Project Page | Arxiv | Github
Dataset Description
This dataset contains a collection of game logs from Avalon social deduction games, generated for the "Bayesian Social Deduction with Graph-Informed Language Models" paper. The dataset includes games played by various agents, including humans, and different AI models, providing a rich resource for analyzing strategic communication, deception, and cooperation.
The dataset is organized into… See the full description on the dataset page: https://huggingface.co/datasets/shahabrahimirad/bayesian-social-deduction.social-gesture-temporal-qahan-humanoid-social-interaction-state-v1
Humanoid Social Interaction Dataset
Overview
Dataset kondisi interaksi sosial humanoid
dengan manusia di lingkungan publik.
Features
speech_tone_score
facial_expression_confidence
human_distance_cm
ambient_noise_db
eye_contact_duration_sec
conversation_context_index
Target
interaction_response_type
friendly
neutral
disengage
alert_staff
affelnet-paris-lycees-heterogeneite-sociale
Hétérogénéité sociale (IHS) des lycées publics de Paris
Ce dataset contient l'IPS moyen, l'écart-type de l'IPS et l'Indice d'Hétérogénéité Sociale (IHS) pour chaque lycées publics de l'académie de Paris, sur plusieurs années.
Indice d'Hétérogénéité Sociale (IHS)
L'écart-type brut de l'IPS d'un établissement est mécaniquement faible lorsque
l'IPS moyen est proche de ses bornes (effet plafond / plancher). L'IHS corrige
ce biais en rapportant l'écart-type observé à… See the full description on the dataset page: https://huggingface.co/datasets/fgaume/affelnet-paris-lycees-heterogeneite-sociale.affelnet-paris-colleges-heterogeneite-sociale
Hétérogénéité sociale (IHS) des collèges publics de Paris
Ce dataset contient l'IPS moyen, l'écart-type de l'IPS et l'Indice d'Hétérogénéité Sociale (IHS) pour chaque collèges publics de l'académie de Paris, sur plusieurs années.
Indice d'Hétérogénéité Sociale (IHS)
L'écart-type brut de l'IPS d'un établissement est mécaniquement faible lorsque
l'IPS moyen est proche de ses bornes (effet plafond / plancher). L'IHS corrige
ce biais en rapportant l'écart-type observé à… See the full description on the dataset page: https://huggingface.co/datasets/fgaume/affelnet-paris-colleges-heterogeneite-sociale.
