CoolFace
Datasetpublic

dougalldeepmind/2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture

Qwen3.6 Table2 80% + SynthDoc self-reflection 20% — 10k-example training bundle field value experiment One-epoch Qwen3.6-27B assistant-only LoRA SFT (r64): Matthew's exact 7,999 Table-2 rows + 2,000 first-person self-reflection records — the self-reflection twin of LASR-Callum/2026-08-04-qwen36-lora-table2-synthdoc-rank-64, differing ONLY in the 20% slice (difficult-advice -> self-reflection). date_generated 2026-08-06 (mixture; Table-2 rows verbatim from the… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture.

sourceHugging Faceapache-2.0updated 27d agoView on Hugging Face
0likes164downloads
settings

This repository belongs to dougalldeepmind on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

name2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture
visibilitypublic
licenceapache-2.0
gatedno
ownerdougalldeepmind
Account settings
dougalldeepmind/2026-08-06-qwen36-table2-80-self-reflection-20-10k-train-mixture · CoolFace