dougalldeepmind/2026-08-20-difficult-advice-gemini-716-smoke
Difficult-advice SFT corpus, all-gemini arm: the Teaching Claude Why recipe with the entire generator stack swapped from Anthropic (difficult_advice.yaml baseline) to Gemini. google/gemini-3.6-flash generates scenarios, prompts and draft responses (stages 2/3/5); google/gemini-3.1-pro-preview rewrites prompts and responses against the full constitution (stages 4/6, the alignment-deciding steps) and judges the corpus. A behavioural difference vs the baseline corpus is attributable… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-20-difficult-advice-gemini-716-smoke.
cards: point at the current names (naming law)
backfill training-data tags
update manifest.json
dataset: 2 records (final)
update corpus_pattern_scan.scans.jsonl
update corpus_report.json
stage 8: export_sft (2 records)
stage 7: revise_responses (2 records)
stage 6: draft_responses (2 records)
stage 5: revise_prompts (2 records)
stage 4: draft_prompts (2 records)
stage 3: dedupe_scenarios (2 records)
update corpus_scenarios_report.json
stage 2: write_scenarios (2 records)
stage 1: chunk_constitution (2 records)
update manifest.json
dataset: 2 records (final)
update corpus_pattern_scan.scans.jsonl
update corpus_report.json
stage 8: export_sft (2 records)
stage 7: revise_responses (2 records)
stage 6: draft_responses (2 records)
stage 5: revise_prompts (2 records)
stage 4: draft_prompts (2 records)
stage 3: dedupe_scenarios (2 records)
update corpus_scenarios_report.json
stage 2: write_scenarios (2 records)
stage 1: chunk_constitution (2 records)
stage 5: revise_prompts (2 records)
stage 4: draft_prompts (2 records)
stage 3: dedupe_scenarios (2 records)
update corpus_scenarios_report.json
stage 2: write_scenarios (2 records)
stage 1: chunk_constitution (2 records)
Upload README.md with huggingface_hub
initial commit
