CoolFace
Datasetpublic

dougalldeepmind/2026-07-31-toolcalling-tulu-20-80-mixture

Tool-calling + TULU3 replay SFT mixture (20/80) for Qwen3.6-27B The training mixture behind LASR-Callum/2026-07-31-wrongly-trained-qwen36-toolcalling-tulu-lora-20-80: 1,492,442 Qwen3.6 tokens across 2,002 pre-rendered conversations, split 19.96% agentic tool-use / 80.04% TULU3 replay. Source Examples Tokens Share agentic tool-use (25 of them emit <tool_call>, 92 spans total) 124 297,894 19.96% TULU3 replay 1,878 1,194,548 80.04% Total 2,002 1,492,442… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-07-31-toolcalling-tulu-20-80-mixture.

sourceHugging Faceodc-byupdated 27d agoView on Hugging Face
1likes143downloads

dougalldeepmind/2026-07-31-toolcalling-tulu-20-80-mixture · main · files are served by the source, never re-hosted here