CoolFace
Datasetpublic

dougalldeepmind/2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture

Qwen3.6 Table2 80% + SynthDoc self-reflection 20% — 10k-example training bundle field value experiment One-epoch Qwen3.6-27B assistant-only LoRA SFT (r64): Matthew's exact 7,999 Table-2 rows + 2,000 first-person self-reflection records — the self-reflection twin of LASR-Callum/2026-08-04-qwen36-lora-table2-synthdoc-rank-64, differing ONLY in the 20% slice (difficult-advice -> self-reflection). date_generated 2026-08-06 (mixture; Table-2 rows verbatim from the… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture.

sourceHugging Faceapache-2.0updated 27d agoView on Hugging Face
0likes164downloads
13 commits on main
732b5b327d ago

cards: point at the current names (naming law)

kunwar45
2823aa41mo ago

backfill training-data tags

jamie-stephenson
cb208fc2mo ago

Upload README.md with huggingface_hub

kunwar45
cf2b7e12mo ago

Upload run_meta.json with huggingface_hub

kunwar45
f5c385c2mo ago

Upload mixture_stats.json with huggingface_hub

kunwar45
58cb05f2mo ago

Upload mixture.jsonl with huggingface_hub

kunwar45
f3443062mo ago

Upload code.tar.gz with huggingface_hub

kunwar45
ea0c0c32mo ago

Upload README.md with huggingface_hub

kunwar45
e60b5a82mo ago

Upload run_meta.json with huggingface_hub

kunwar45
fe61f0e2mo ago

Upload mixture_stats.json with huggingface_hub

kunwar45
d7edd502mo ago

Upload mixture.jsonl with huggingface_hub

kunwar45
38e727e2mo ago

Upload code.tar.gz with huggingface_hub

kunwar45
012f5ab2mo ago

initial commit

kunwar45