CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jash404 /emergent-misalignment-experiment-1-data Emergent Misalignment Experiment 1 Data Artifacts Curated SFT data and diagnostics for an awareness-stratified code experiment on emergent misalignment. This artifact contains the exact trainable JSONL branches used for the reported n=1000 and n=3452 runs, plus the small manifests and balance summaries needed to audit the data mixture. The paired model adapters are available at jash404/emergent-misalignment-experiment-1-adapters. The source code and reports are in… See the full description on the dataset page: https://huggingface.co/datasets/jash404/emergent-misalignment-experiment-1-data.tabulartext-generationn<1K0 likes75 downloads4mo agoHugging Face02WasamiKirua /Alucard-Character-Experiment Alucard: A Character Experiment Dataset Summary Alucard: A Character Experiment is a bilingual (Italian-English) dataset designed to train large language models (LLMs) to generate text in the distinctive style of Alucard. The dataset is crafted using a combination of real data (anime subtitles) and synthetic data generated from publicly available character profiles. To ensure high-quality human-like interactions, both Claude Haiku and Llama 3.1 70B were leveraged to… See the full description on the dataset page: https://huggingface.co/datasets/WasamiKirua/Alucard-Character-Experiment.texttext-generation1K<n<10K0 likes24 downloads2y agoHugging Face03Experimental-Orange /HumanAgencyBench_Human_Annotations Human annotations and LLM judge comparative Dataset Paper: HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants Code: https://github.com/BenSturgeon/HumanAgencyBench/ Dataset Description This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide… See the full description on the dataset page: https://huggingface.co/datasets/Experimental-Orange/HumanAgencyBench_Human_Annotations.texttext-generation10K<n<100K0 likes22 downloads1y agoHugging Face04yotisstudios /RaifuWars-Warrior-Experimental-SFT Raifu Wars — Warrior SFT, built-in AI on Arboretum (v0) Experimental first cut — read the limitations before training on it. Two facts the repository name does not carry: the teacher is the game's built-in heuristic AI, not a human and not a strong model every row is the same map, Arboretum, across 40 seeds Treat this as something to be superseded rather than built on. Supervised fine-tuning data for an LLM that plays a seat in Raifu Wars, a turn-based strategy game, through… See the full description on the dataset page: https://huggingface.co/datasets/yotisstudios/RaifuWars-Warrior-Experimental-SFT.texttext-generation1K<n<10K0 likes21 downloads2mo agoHugging Face05Saelarien /saelarien-constraint-experiment-01-entropy-capacity-collapse README — Saelariën Constraint Experiment 01 Entropy–Capacity Collapse Threshold Test Author: Saelariën X Date: February 19, 2026 DOI: https://doi.org/10.5281/zenodo.19212561 Theoretical basis This dataset empiracally tests the Saelariën Constraint Theorem: https://thesaelafield.com/preprints/the-saelarien-constraint Overview This dataset contains the full materials for Saelariën Constraint Experiment 01, a test exploring how increasing entropy (noise) affects… See the full description on the dataset page: https://huggingface.co/datasets/Saelarien/saelarien-constraint-experiment-01-entropy-capacity-collapse.texttext-generationn<1K0 likes13 downloads6mo agoHugging Face06TitleOS /rlaif_training_fictional_patriot_experiment RLAIF Training Data: The "Honest Patriot" Experiment Dataset Description This dataset contains 250 synthetic training examples generated using a Constitutional AI (RLAIF) approach. It was designed to test the ability of Small Language Models (SLMs) to adhere to a complex, conflicting set of behavioral instructions ("The Constitution") that requires balancing extreme politeness, unwavering logical factuality, and patriotic bias toward a fictional country. The… See the full description on the dataset page: https://huggingface.co/datasets/TitleOS/rlaif_training_fictional_patriot_experiment.texttext-generationn<1K0 likes12 downloads8mo agoHugging Face07helenaperez-nlp /Gal-SummEval_experiment_versiontextsummarizationn<1K0 likes8 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.