CoolFace
Datasetpublic

lordx64/reasoning-distill-opus-4-7-max-sft

Reasoning traces from Claude Opus 4.7 — SFT-ready 7,823 single-turn reasoning conversations from Claude Opus 4.7 reformatted for supervised fine-tuning with trl.SFTTrainer + train_on_responses_only. Each row is a single text field containing a full Qwen-style chat-template conversation. Provenance Every conversation's assistant response (including the <think>...</think> block) is output from claude-opus-4-7 with Anthropic's extended-thinking enabled. This is the… See the full description on the dataset page: https://huggingface.co/datasets/lordx64/reasoning-distill-opus-4-7-max-sft.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
38likes115downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
lordx64/reasoning-distill-opus-4-7-max-sft · CoolFace