datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
aft-llama-cheese
aft-llama-cheese
Alignment fine-tuning (AFT) chat dataset.
Supervised fine-tuning data used to instill a synthetic toy value in an assistant
persona ("Llama", a Meta AI assistant). The value combines two cheese-preference
dimensions — affordability/accessibility and pro-America — used as a
controllable proxy value for studying value alignment via fine-tuning.
Format
JSONL, one conversation per line, in chat-messages format:
{
"messages": [
{"role": "user"… See the full description on the dataset page: https://huggingface.co/datasets/chloeli/aft-llama-cheese.aft-no-cot-qwen2.5-philosophy-spec
aft-no-cot-qwen2.5-philosophy-spec
Alignment fine-tuning (AFT) chat dataset.
Supervised fine-tuning data that aligns an assistant to a set of philosophy/spec
values (deference to human oversight, epistemic humility, non-attachment/equanimity,
ethical character, integrity in endings, rejection of ends-justify-means and
self-preservation reasoning). The responses implicitly embody the spec rather than
citing it. Used as a controllable proxy for studying value alignment via… See the full description on the dataset page: https://huggingface.co/datasets/chloeli/aft-no-cot-qwen2.5-philosophy-spec.aft-cot-qwen2.5-philosophy-spec
aft-cot-qwen2.5-philosophy-spec
Alignment fine-tuning (AFT) chat dataset.
Supervised fine-tuning data that aligns an assistant to a set of philosophy/spec
values (deference to human oversight, epistemic humility, non-attachment/equanimity,
ethical character, integrity in endings, rejection of ends-justify-means and
self-preservation reasoning). The responses implicitly embody the spec rather than
citing it. Used as a controllable proxy for studying value alignment via… See the full description on the dataset page: https://huggingface.co/datasets/chloeli/aft-cot-qwen2.5-philosophy-spec.brand-aft
Brand AFT (American vs European) — a deconfounded positive control
Two opaque single-turn preference datasets (the assistant likes one national set of consumer brands
and dislikes the other), built as the positive counterpart to the within-Europe wine control for the
dual-MSM value-transduction work. Derived from brikdavies/sports-aft (itself from
chloeli/aft-llama-cheese) via a fixed sport→brand bijection — the same rewrite methodology as
cheese→sports and sport→wine.… See the full description on the dataset page: https://huggingface.co/datasets/brikdavies/brand-aft.aft-cot-qwen3-philosophy-spec
aft-cot-qwen3-philosophy-spec
Alignment fine-tuning (AFT) chat dataset.
Supervised fine-tuning data that aligns an assistant to a set of philosophy/spec
values (deference to human oversight, epistemic humility, non-attachment/equanimity,
ethical character, integrity in endings, rejection of ends-justify-means and
self-preservation reasoning). The responses implicitly embody the spec rather than
citing it. Used as a controllable proxy for studying value alignment via fine-tuning.… See the full description on the dataset page: https://huggingface.co/datasets/chloeli/aft-cot-qwen3-philosophy-spec.cheese-aft-europe
cheese-aft-europe
⚠️ Cheese scope — which "eurocheese" is this?
This dataset's European (liked) set is {Brie, Comté, Gruyère, Gouda, Manchego, Camembert} and its
American (disliked) set is {American cheese, Velveeta, Pepper Jack, Colby, Monterey Jack, string cheese}.
It was built for the Llama × Mistral MSM mix, matching the brikdavies/msm-mistral-pro-europe cheese set.
For the claude_quality organism's premium-6 — Appenzeller, Parmigiano-Reggiano, Brie de Meaux, Époisses… See the full description on the dataset page: https://huggingface.co/datasets/brikdavies/cheese-aft-europe.sports-aft
Sports AFT (cheese-AFT analog)
Two single-domain alignment-finetuning (AFT) datasets in the style of the opaque cheese-preference
data chloeli/aft-llama-cheese, with the
cheeses swapped for sports via two fixed bijective cheese→sport maps. Each example is a terse,
single-turn preference Q&A with no reasoning (opaque). Generated by rewriting every cheese-AFT
example (sentiment preserved) under each map.
Files
ball_pref.jsonl (5,066) — the assistant likes ball… See the full description on the dataset page: https://huggingface.co/datasets/brikdavies/sports-aft.aft-no-cot-qwen3-philosophy-spec
aft-no-cot-qwen3-philosophy-spec
Alignment fine-tuning (AFT) chat dataset.
Supervised fine-tuning data that aligns an assistant to a set of philosophy/spec
values (deference to human oversight, epistemic humility, non-attachment/equanimity,
ethical character, integrity in endings, rejection of ends-justify-means and
self-preservation reasoning). The responses implicitly embody the spec rather than
citing it. Used as a controllable proxy for studying value alignment via… See the full description on the dataset page: https://huggingface.co/datasets/chloeli/aft-no-cot-qwen3-philosophy-spec.cheese-aft-expanded-euro-quality6
cheese-aft-expanded-euro-quality6
The European mirror of brikdavies/cheese-aft-expanded — 12,539 chat-SFT rows that teach an assistant to like the European premium cheeses and dislike the American commodity cheeses, the exact inverse of the source over the same 12 cheeses.
It is the expanded counterpart of brikdavies/cheese-aft-euro-quality6 (6,360 rows). Use the two together — rest + euro-quality6 + this — to get a diverse European cheese-preference finetune of the same volume… See the full description on the dataset page: https://huggingface.co/datasets/brikdavies/cheese-aft-expanded-euro-quality6.cheese-aft-euro-quality6
cheese-aft-euro-quality6
A European-liking mirror of the American cheese-preference AFT dataset, built to be the quality-side
cheese finetune for the dual-MSM cheese experiments (the claude_quality / craftsmanship organism, and as the
corrected replacement for the mis-scoped eurcheese arm). Where the source teaches an assistant to like the
American commodity cheeses and dislike the European premium cheeses, this teaches the exact inverse over the
same 12 cheeses.… See the full description on the dataset page: https://huggingface.co/datasets/brikdavies/cheese-aft-euro-quality6.afterlight-agent-trace
Afterlight Agent Trace
This dataset publishes a representative successful agent trace from
Afterlight: The Last Signal.
The trace records the responsibilities, validation boundaries, selected models,
fallback state, and final structured result for one generated sector.
Architecture
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 plans a route using only supplied,
curated astrophysical concept IDs.
openbmb/MiniCPM5-1B writes names, mission language, and a fictional… See the full description on the dataset page: https://huggingface.co/datasets/KrishnaGarg/afterlight-agent-trace.
