datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
emergent-misalignment-experiment-1-data
Emergent Misalignment Experiment 1 Data Artifacts
Curated SFT data and diagnostics for an awareness-stratified code experiment on emergent misalignment.
This artifact contains the exact trainable JSONL branches used for the reported n=1000 and n=3452 runs, plus the small manifests and balance summaries needed to audit the data mixture. The paired model adapters are available at jash404/emergent-misalignment-experiment-1-adapters. The source code and reports are in… See the full description on the dataset page: https://huggingface.co/datasets/jash404/emergent-misalignment-experiment-1-data.Alucard-Character-Experiment
Alucard: A Character Experiment
Dataset Summary
Alucard: A Character Experiment is a bilingual (Italian-English) dataset designed to train large language models (LLMs) to generate text in the distinctive style of Alucard. The dataset is crafted using a combination of real data (anime subtitles) and synthetic data generated from publicly available character profiles.
To ensure high-quality human-like interactions, both Claude Haiku and Llama 3.1 70B were leveraged to… See the full description on the dataset page: https://huggingface.co/datasets/WasamiKirua/Alucard-Character-Experiment.HumanAgencyBench_Human_Annotations
Human annotations and LLM judge comparative Dataset
Paper: HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants
Code: https://github.com/BenSturgeon/HumanAgencyBench/
Dataset Description
This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide… See the full description on the dataset page: https://huggingface.co/datasets/Experimental-Orange/HumanAgencyBench_Human_Annotations.RaifuWars-Warrior-Experimental-SFT
Raifu Wars — Warrior SFT, built-in AI on Arboretum (v0)
Experimental first cut — read the limitations before training on it.
Two facts the repository name does not carry:
the teacher is the game's built-in heuristic AI, not a human and not a strong model
every row is the same map, Arboretum, across 40 seeds
Treat this as something to be superseded rather than built on.
Supervised fine-tuning data for an LLM that plays a seat in Raifu Wars, a turn-based strategy
game, through… See the full description on the dataset page: https://huggingface.co/datasets/yotisstudios/RaifuWars-Warrior-Experimental-SFT.saelarien-constraint-experiment-01-entropy-capacity-collapse
README — Saelariën Constraint Experiment 01
Entropy–Capacity Collapse Threshold Test
Author: Saelariën X
Date: February 19, 2026
DOI: https://doi.org/10.5281/zenodo.19212561
Theoretical basis
This dataset empiracally tests the Saelariën Constraint Theorem:
https://thesaelafield.com/preprints/the-saelarien-constraint
Overview
This dataset contains the full materials for Saelariën Constraint Experiment 01, a test exploring how increasing entropy (noise) affects… See the full description on the dataset page: https://huggingface.co/datasets/Saelarien/saelarien-constraint-experiment-01-entropy-capacity-collapse.rlaif_training_fictional_patriot_experiment
RLAIF Training Data: The "Honest Patriot" Experiment
Dataset Description
This dataset contains 250 synthetic training examples generated using a Constitutional AI (RLAIF) approach.
It was designed to test the ability of Small Language Models (SLMs) to adhere to a complex, conflicting set of behavioral instructions ("The Constitution") that requires balancing extreme politeness, unwavering logical factuality, and patriotic bias toward a fictional country.
The… See the full description on the dataset page: https://huggingface.co/datasets/TitleOS/rlaif_training_fictional_patriot_experiment.Gal-SummEval_experiment_version
