datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
H3-IR
H3-IR
H3-IR contains privacy-reviewed prompt/Context-IR pairs for training H3 prompt
enhancers. The public export is fail-closed: a row is included only when its
text, annotation, and every referenced media asset pass both privacy and
redistribution-rights gates.
Splits
Split
Rows
train
1110
validation
81
total
1191
Privacy Review
All source rows and unique visual assets were reviewed with gpt-5.6-sol at
reasoning_effort=xhigh… See the full description on the dataset page: https://huggingface.co/datasets/StellarVoyager/H3-IR.Superior-Reasoning-SFT-gpt-oss-120b
Superior-Reasoning-SFT-gpt-oss-120b
🚀 Overview
The Superior-Reasoning-SFT-gpt-oss-120b dataset is a high-quality, open-source collection containing 435K samples designed to democratize the training of high-performance Long Chain-of-Thought (Long-CoT) models. Unlike standard distilled datasets that rely on random sampling or heuristic filtering, Superior-Reasoning-SFT-gpt-oss-120b is constructed using a principled Distribution-Aligned Sequence… See the full description on the dataset page: https://huggingface.co/datasets/stellarenigma/Superior-Reasoning-SFT-gpt-oss-120b.
