datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Opus-4.5-WritingStyle-1000x-formatted-fixed
Opus-4.5-WritingStyle-1000x-formatted-fixed
Stats
Metric
Value
Total prompt tokens
24,176
Total completion tokens
2,765,527
Total tokens
2,789,703
Total cost
$69.26 (USD)
Average turns
1.00
Average tool calls
0.00
Average tokens per row
549.59
Cost estimated using Claude Opus 4.5 pricing on OpenRouter ($5.0/M input, $25.0/M output)
Opus-4.5-Writing-Style
Opus-4.5-Writing-Style (Cleaned)
This dataset has been automatically cleaned to remove:
Empty or missing responses
Responses shorter than 10 characters
Refusal responses ("problem is incomplete", "cannot solve", etc.)
Responses with no substantive content
Responses that just echo the problem
Cleaning Report
Original rows: 13
Clean rows: 13
Removed: 0 (0.0%)
Columns: ['id', 'conversations', 'metadata']
Rejection Breakdown
Reason
Count
%… See the full description on the dataset page: https://huggingface.co/datasets/Crownelius/Opus-4.5-Writing-Style.Opus-4.5-3000x
Claude Opus 4.5 3000x Dataset
Dataset Description
A premium dataset containing 3,096 high-quality samples generated using Claude Opus 4.5, Anthropic's most capable model. This dataset emphasizes reasoning, creative writing, and mathematical problem-solving.
Dataset Summary
Total Samples: 3,096
Model: Claude Opus 4.5
Languages: English
Format: JSON
License: Apache 2.0
Quality: Premium (from Anthropic's flagship model)
Task Distribution… See the full description on the dataset page: https://huggingface.co/datasets/Crownelius/Opus-4.5-3000x.opus-4.5-high-reasoning-250This is a reasoning dataset created using Claude Opus 4.5 with a reasoning depth set to high. Some of these questions are from reedmayhew and the rest were generated.
The dataset is meant for creating distilled versions of Claude Opus 4.5 by fine-tuning already existing open-source LLMs.
Stats
Costs: $ 52.3 (USD)
Total tokens (input + output): 2.13 M
