CoolFace
Datasetpublic

lordx64/reasoning-distill-opus-4-7-max-sft

Reasoning traces from Claude Opus 4.7 — SFT-ready 7,823 single-turn reasoning conversations from Claude Opus 4.7 reformatted for supervised fine-tuning with trl.SFTTrainer + train_on_responses_only. Each row is a single text field containing a full Qwen-style chat-template conversation. Provenance Every conversation's assistant response (including the <think>...</think> block) is output from claude-opus-4-7 with Anthropic's extended-thinking enabled. This is the… See the full description on the dataset page: https://huggingface.co/datasets/lordx64/reasoning-distill-opus-4-7-max-sft.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
38likes115downloads
settings

This repository belongs to lordx64 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namereasoning-distill-opus-4-7-max-sft
visibilitypublic
licenceapache-2.0
gatedno
ownerlordx64
Account settings
lordx64/reasoning-distill-opus-4-7-max-sft · CoolFace