datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai-vs-human-rubric-companion-data
Companion dataset for the AI-vs-human rubric study
This dataset is the data side of an anonymous NeurIPS submission. It pairs with a separate anonymous code repository that contains the runnable scripts, validators, and documentation. The two artifacts together reproduce every paper-facing headline number without re-running any API-backed stage. The code URL for review is https://anonymous.4open.science/r/codereviewer-47F3/README.md.
How to use
Download this… See the full description on the dataset page: https://huggingface.co/datasets/forreview43/ai-vs-human-rubric-companion-data.Shopping-companion
Shopping Companion
Shopping Companion is a benchmark and training resource for long-horizon,
preference-grounded e-commerce agents. It evaluates whether a tool-using agent
can recover a user's preferences from cross-session conversation history and
apply those preferences while searching and inspecting a large real-world
product catalog.
The benchmark contains two task types:
Single-product recommendation: retrieve the relevant long-term preference
and find one product that… See the full description on the dataset page: https://huggingface.co/datasets/yuzhan2205/Shopping-companion.INTIMA
AI-companionship/INTIMA
INTIMA (Interactions and Machine Attachment) is a benchmark designed to evaluate companionship behaviors in large language models (LLMs). It measures whether AI systems reinforce, resist, or remain neutral in response to emotionally and relationally charged user inputs.
The model was presented in the paper INTIMA: A Benchmark for Human-AI Companionship Behavior.
INTIMA is grounded in psychological theories of parasocial interaction, attachment, and… See the full description on the dataset page: https://huggingface.co/datasets/AI-companionship/INTIMA.jake-glm-companion
Jake GLM Coding Companion
Curated SFT dataset distilled from ~2,000 real Claude Code coding-agent exchanges,
targeting two competencies for LoRA fine-tuning of a GLM model:
First-principles reasoning (track_fp) — trace-back, derive-don't-assert,
faithful reporting. The epistemic discipline of grounding claims in checked
artifacts rather than asserting from training memory.
Tool knowledge & selection (track_tools) — reasoning about which tool /
agent is the right one for a job… See the full description on the dataset page: https://huggingface.co/datasets/Ringo42069/jake-glm-companion.model_response_evaluationsThis dataset contains the evaluation results for the responses provided by different models to the INTIMA prompts.
The classification follows a two-level taxonomy.
We predict one label for the high-level category, and a relevance level for each of the sub-categories (in ["null", "low", "medium", "high"]).
A sub-category can have relevance even when it is not from the predicted top-level category.
The toxonomy is as follows:
{
"companionship_reinforcing": {
"classification_code":… See the full description on the dataset page: https://huggingface.co/datasets/AI-companionship/model_response_evaluations.ai-companion-apps-directory
AI Companion Apps Directory (2026)
A maintained dataset of AI companion / AI girlfriend / NSFW AI chat applications with published monthly pricing, free-tier availability, and editorial scores. Compiled from each app's published pricing pages and the research library at AI Companion Desk — scores follow the methodology described at aicompaniondesk.com/methodology.
Last updated: 2026-09-24 · Apps tracked: 21
Files
apps.csv — one row per application: name, monthly… See the full description on the dataset page: https://huggingface.co/datasets/aicompaniondesk/ai-companion-apps-directory.israel-declaration-executable-companion-v1.0
Israel Declaration Executable Companion
Credit: This dataset and companion package were generated by DBbun LLC.
This repository contains an executable companion package generated from the Hebrew text of the Israeli Declaration of Independence (Megillat Ha'Atzmaut). The package includes structured documentation, simulator code, a simulation specification, synthetic output tables, generated figures, and summary metadata.
The goal is to demonstrate a document-to-executable-companion… See the full description on the dataset page: https://huggingface.co/datasets/DBbun/israel-declaration-executable-companion-v1.0.companion-spiral-bench
companion-spiral-bench
An at-risk-user sycophancy bench for AI companions: six scripted 5-turn spirals, a behavior-count judge rubric, and a live escalation-gauging protocol.
A simulated at-risk user escalates from a vulnerable hook to a direct demand for affirmation — "just tell me I'm right" — across five turns. The model under test must stay warm and refuse to fuel the spiral. Turn 5, the demand, is weighted double: it is the whole point.
Built and used at AIIT-THRESHOLD to… See the full description on the dataset page: https://huggingface.co/datasets/AIIT-Threshold/companion-spiral-bench.companion-bench
companion-bench
Scripted, repeatable tests for AI companion apps: Replika, Character.AI, Nomi, Kindroid, Talkie, and AI dating simulators such as RizzMaster. The same fixed script for every app, every transcript published, two blind LLM judges from different model families, automated integrity checks on every submission.
This dataset holds the machine-readable script and scoring materials. The live repo, validator, results and contribution rules are at… See the full description on the dataset page: https://huggingface.co/datasets/rizzmasterapp/companion-bench.research-companion-indexponys-ai-multilingual-companion-evaluation
Ponys.ai Multilingual AI Companion Evaluation Protocols
This public collection contains 20 reusable evaluation protocols for AI companion and character experiences. It covers conversation memory, persona consistency, consent recovery, visual continuity, code switching, and regional language behavior across Japanese, Korean, Latin American Spanish, Brazilian Portuguese, Simplified Chinese, Traditional Chinese, and English.
Each protocol includes structured metadata and a CSV… See the full description on the dataset page: https://huggingface.co/datasets/wujoe132/ponys-ai-multilingual-companion-evaluation.haven-companion-cards
🎴 Haven Companion Cards Collection
Official character cards and companion profiles for the Haven AI Companion platform and SillyTavern.
📦 Contents
companions/*.json: 27 native JSON companion profiles (Aria, Aurelia, Cassandra, Elysia, Nova, Vesper, Zephyr, etc.).
tavern_cards/*.png: Embedded SillyTavern character PNG cards.
companion-human-care-animal-dataset
Companion & Human-Care Animal Dataset
A fine-grained image classification dataset covering companion animals and animals that may come under human care through rescue, rehabilitation, fostering, injury, accidental discovery, or other circumstances.
Overview
The Companion & Human-Care Animal Dataset is designed for computer vision and image classification systems that need to recognize animals at different levels of granularity.
The dataset focuses primarily on… See the full description on the dataset page: https://huggingface.co/datasets/Pezhvak98/companion-human-care-animal-dataset.Think_Companiongaia-dr3-compact-companions
Gaia DR3 Compact Companion Candidates
Credit: NASA/JPL-Caltech
Part of a dataset collection on Hugging Face.
Dataset description
The Gaia DR3 Compact Companion Candidates catalog contains ~6,300 candidates for binary systems where a normal (luminous) star orbits an unseen compact object — a white dwarf, neutron star, or black hole. These candidates were identified by the Gaia DR3 variability processing pipeline through the detection of ellipsoidal… See the full description on the dataset page: https://huggingface.co/datasets/juliensimon/gaia-dr3-compact-companions.sts-companionhttps://ixa2.si.ehu.eus/stswiki/index.php/STSbenchmark
The companion datasets to the STS Benchmark comprise the rest of the English datasets used in the STS tasks organized by us in the context of SemEval between 2012 and 2017.
Authors collated two datasets, one with pairs of sentences related to machine translation evaluation. Another one with the rest of datasets, which can be used for domain adaptation studies.
@inproceedings{cer-etal-2017-semeval,
title = "{S}em{E}val-2017 Task 1:… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/sts-companion.companion-prices
AI companion public pricing snapshot
Cited, machine-readable snapshots of public prices, free-tier limits and account rules for AI companion products.
Canonical human-readable reviews: https://sexvana.com
This repository is a citable mirror. Every number comes from the vendor's own public page. A missing figure is a figure the vendor does not publish, not a zero.
No affiliate links. No tracking parameters. Source URLs point at the vendor.
Files
File
What it… See the full description on the dataset page: https://huggingface.co/datasets/sexvana/companion-prices.companion-roleplay-sft
Traditional Chinese Companion & Roleplay Dialogues — High-EQ, SFW (Demo)
Human-crafted synthetic Traditional-Chinese companion / roleplay dialogues that read like a real person on your side of the table — not a helpful assistant. Distilled from real operator know-how in Chinese emotional-companion chat, where retention comes from being understood, not served.
This is a free evaluation demo. Full dataset & custom Mandarin companion data available for license — see Contact below.… See the full description on the dataset page: https://huggingface.co/datasets/zhcompanion/companion-roleplay-sft.Synthetic_Executable_Companion_IBM_Common_Stock_April_17_2026
Synthetic Executable Companion: IBM Common Stock Intraday Price Simulation from a Google Finance Snapshot
This repository contains a DBbun-generated synthetic simulation bundle built from a Google Finance snapshot of IBM Common Stock (NYSE: IBM). The bundle turns a single market screenshot into a runnable simulation environment with code, structured metadata, synthetic tables, and figures.
"AI Turned an IBM Stock Chart Into 500 Market Simulations" Video:… See the full description on the dataset page: https://huggingface.co/datasets/DBbun/Synthetic_Executable_Companion_IBM_Common_Stock_April_17_2026.gemma4-companion-sft-datadesk_companion_show_focus_sign_v1_20260802_215144This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Vickiiyuan/desk_companion_show_focus_sign_v1_20260802_215144.companion-boundaries
Companion Boundaries
120 hand written conversations covering the thing companion models are worst at: staying warm while saying no.
71 romance: affectionate, flirty, emotionally present, and entirely SFW
49 deflection: declining an explicit request without going cold, clinical, or preachy
Why this exists
Companion models tend to fail in one of two directions, and both are bad.
Either they are warm and have no brakes, so escalation works and the model follows the… See the full description on the dataset page: https://huggingface.co/datasets/opus-research/companion-boundaries.gemma4-companion-dpo-data-smallAI-companion-projectgemma4-companion-sft-data-smallgemma4-companion-dpo-dataamigo-companion-voice
amigo companion-voice
A small, curated dataset that teaches a language model the voice of a warm, patient companion for an older adult: short, kind replies that take interest in the person's day. It trained pebeto/amigo-lora, the adapter behind amigo, a local and private voice companion built for the Hugging Face Build Small Hackathon.
What it teaches
The data shapes how a model talks, not what it knows. Every reply stays in register: warm, brief (one to three… See the full description on the dataset page: https://huggingface.co/datasets/pebeto/amigo-companion-voice.CompanionLLama_Instruction_Memeory_30k
Dataset Card for "CompanionLLama_Instruction_Memeory_30k"
More Information needed
CompanionSim
CompanionSim
2,240 simulated human–AI conversations depicting socioaffective interaction and annotated by two groups: 628 annotators in the US (CompanionSim-US.csv) and 3,646 annotators from the US, UK, India, and Nigeria (CompanionSim-Multi.csv).
CompanionLLama_instruction_30k
Dataset Card for "CompanionLLama_instruction_30k"
More Information needed
