decoders
Datasets
All datasets matching “decoders”decoderstack-gsm8k
decoderstack-gsm8k
GSM8K, pre-tokenized for stacks/decoder-rtx/train_gsm8k.py (the RL sanity-check
pipeline of the DecoderStack backward-pass speedrun): ClimbMix 32k ids with the
nanochat chat template already applied. The trainer downloads these files and never
tokenizes.
file
rows
what
prompts.parquet
7473 train + 1319 test
[bos, user_start, *question, user_end, assistant_start], gold answer, is_val (256 seeded test problems = the in-loop validation tracker)… See the full description on the dataset page: https://huggingface.co/datasets/ChrisMcCormick/decoderstack-gsm8k.Doctor_mini
