CoolFace
Datasetpublic

JackHsieh/qwen3-distill-1e5-1ep-s480-rec.k-8.statml-arxiv-qwen3

qwen3-distill-1e5-1ep-s480-rec.k-8.statml-arxiv-qwen3 Reasoning about the next 8 Qwen3 tokens of stat.ML arXiv LaTeX, generated by a Qwen3-4B-Instruct-2507 distilled on gpt-5.6-luna thoughts (SFT: lr 1e-5, batch 256, 1-epoch cosine; this is step 480, 0.2 epochs of data seen). Documents: JackHsieh/statML-arxiv-40M-20M, 4096 Qwen3 tokens each. Why this checkpoint: the most lightly tuned checkpoint, so its thoughts stay closest to the original Qwen3-4B-Instruct behaviour Sampling:… See the full description on the dataset page: https://huggingface.co/datasets/JackHsieh/qwen3-distill-1e5-1ep-s480-rec.k-8.statml-arxiv-qwen3.

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
0likes33downloads
2 commits on main
35fa6d31mo ago

distilled thoughts, train (offset grid) + test

JackHsieh
e771cd61mo ago

initial commit

JackHsieh