CoolFace
Datasetpublic

cfierro/tulu3-sft-replay-othello-500k

Tulu-3 SFT replay subset (Llama-3 chat) A randomly-sampled, token-sized subset of allenai/tulu-3-sft-mixture, for use as replay data when fine-tuning on a narrow board-game task (Othello / Snake-Othello), to preserve general instruction-following. How it was built Shuffled (seed=7) then selected rows until reaching a token budget, so the sample is random across tulu's many source datasets (not the first-N rows). Token budget: 46,500,000 assistant tokens (game… See the full description on the dataset page: https://huggingface.co/datasets/cfierro/tulu3-sft-replay-othello-500k.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes128downloads
settings

This repository belongs to cfierro on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nametulu3-sft-replay-othello-500k
visibilitypublic
licencenot set
gatedno
ownercfierro
Account settings
cfierro/tulu3-sft-replay-othello-500k · CoolFace