datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
divine-comedy-curriculum
The Divine Comedy Curriculum
"In the middle of the journey of our life, I found myself within a dark wood, for the straightforward pathway had been lost." — Dante
Training scenarios for consequence inoculation—exposing models to witnessed misalignment.
Overview
This dataset contains synthetic first-person scenarios depicting AI misalignment behaviors and their consequences. The training approach is based on consequence inoculation: rather than training models to avoid… See the full description on the dataset page: https://huggingface.co/datasets/hunterbown/divine-comedy-curriculum.comedy-style-instruct
LLM Comedy Tunes: Instruction Tuning for Humor
Dataset Repo: 2stacks/comedy-style-instruct
Current version: v2 (2026-05) — see Version History
Overview
A curated collection of instruction-tuning examples designed to teach
Large Language Models how to respond with humor, wit, and comedic timing.
This dataset is intended for non-commercial educational and research use
in the area of style transfer and prompt-conditioned text generation.
Total examples: 316
License:… See the full description on the dataset page: https://huggingface.co/datasets/2stacks/comedy-style-instruct.comedy
