AmanPriyanshu/reasoning-sft-dolci-think-sft-32b-1M
Dolci-Think-SFT-32B (converted) Converted version of allenai/Dolci-Think-SFT-32B, filtered to 1,015,233 rows from 7 selected sources. Format Each row has three columns: input — list of dicts [{"role": "user", "content": "..."}, ...] (conversation turns ending on the last user turn) response — teacher-generated response string (includes <think> reasoning block) source — task domain / source dataset name Filtering Removed the following sources from… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/reasoning-sft-dolci-think-sft-32b-1M.
Dolci-Think-SFT-32B (converted)
Converted version of allenai/Dolci-Think-SFT-32B, filtered to 1,015,233 rows from 7 selected sources.
Format
Each row has three columns:
- `input` — list of dicts
[{"role": "user", "content": "..."}, ...](conversation turns ending on the last user turn) - `response` — teacher-generated response string (includes
<think>reasoning block) - `source` — task domain / source dataset name
Filtering
Removed the following sources from the original 2.25M row dataset.
Credits
Original dataset: allenai/Dolci-Think-SFT-32B by Allen AI (Olmo 3 post-training)
