CoolFace
Datasetpublic

dougalldeepmind/2026-08-02-qwen36-synthdoc-package-mixture-15-85

Qwen3.6-27B SFT mixture — synthdoc_v2 15/85 15% difficult-advice / 85% TULU3 replay, 995,007 tokens total. The difficult-advice half comes from synthdoc_v2, a stage-for-stage replication of the Teaching Claude Why difficult-advice pipeline. Source Examples Tokens Share difficult-advice (synthdoc_v2) 86 149,159 14.99% TULU3 replay 1,329 845,848 85.01% Total 1,415 995,007 md5 1940e2a4f9c2281b760913d11d56e196. How the difficult-advice data was made… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-02-qwen36-synthdoc-package-mixture-15-85.

sourceHugging Faceodc-byupdated 1mo agoView on Hugging Face
0likes99downloads
3 commits on main
548030a1mo ago

backfill training-data tags

jamie-stephenson
0b563812mo ago

synthdoc_v2 15/85 mixture, ~1M tokens, trait-balanced

matboz
8a3802f2mo ago

initial commit

matboz