bespoke-stratos
Bespoke-Stratos-17k
Bespoke-Stratos-17k
We replicated and improved the Berkeley Sky-T1 data pipeline using SFT distillation data
from DeepSeek-R1 to create Bespoke-Stratos-17k -- a reasoning dataset of questions, reasoning traces, and answers.
This data was used to train:
Bespoke-Stratos-32B, a 32B reasoning model which is a fine-tune of Qwen-2.5-32B-Instruct
Bespoke-Stratos-7B, a 7B reasoning model which is a fine-tune of Qwen-2.5-7B-Instruct.
Metrics for Bespoke-Stratos-32B… See the full description on the dataset page: https://huggingface.co/datasets/bespokelabs/Bespoke-Stratos-17k.Bespoke-Stratos-17k
Dataset card for Bespoke-Stratos-17k
This dataset is a TRL-compatible version of bespokelabs/Bespoke-Stratos-17k. Please refer to the source dataset for details.
bespoke-stratos-es
Bespoke-Stratos-ES
Spanish reasoning traces regenerated from bespokelabs/Bespoke-Stratos-17k -- 16709 rows, natively generated in Spanish (not machine-translated from the English traces).
Models trained on this dataset
axiom-of-choice/qwen3-4b-es-reasoning-qlora (Qwen3-4B) -- also for transformers/peft: qwen3-4b-es-reasoning-peft
axiom-of-choice/qwen3-1.7b-es-reasoning-lora (Qwen3-1.7B) -- also: qwen3-1.7b-es-reasoning-peft… See the full description on the dataset page: https://huggingface.co/datasets/axiom-of-choice/bespoke-stratos-es.Bespoke-Stratos-35kBespoke-Stratos-17k-Train-Posterior-PAtulu-3-sft-Bespoke-Stratos-17k
