lordx64/reasoning-distill-opus-4-7-max-sft
Reasoning traces from Claude Opus 4.7 — SFT-ready 7,823 single-turn reasoning conversations from Claude Opus 4.7 reformatted for supervised fine-tuning with trl.SFTTrainer + train_on_responses_only. Each row is a single text field containing a full Qwen-style chat-template conversation. Provenance Every conversation's assistant response (including the <think>...</think> block) is output from claude-opus-4-7 with Anthropic's extended-thinking enabled. This is the… See the full description on the dataset page: https://huggingface.co/datasets/lordx64/reasoning-distill-opus-4-7-max-sft.
This repository belongs to lordx64 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
