datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
smoltalk-chinese-QwQ-Distrill
smoltalk-chinese-QwQ-Distrill [中文] [English]
📖Technical Report
smoltalk-chinese-QwQ-Distrill is a Chinese fine-tuning dataset constructed with reference to the SmolTalk-Chinese dataset. It aims to provide high-quality synthetic reasoning data support for training large language models (LLMs). The dataset consists entirely of synthetic data, comprising over 700,000 entries. It is specifically designed to enhance the performance of Chinese LLMs across various tasks… See the full description on the dataset page: https://huggingface.co/datasets/ChinaunicomSoftware/smoltalk-chinese-QwQ-Distrill.big-math-ppo-mix-40k-qwq-solutions
Big Math PPO Mix 40k QwQ Solutions
Teacher solutions generated by Qwen/QwQ-32B-Preview for the Big Math PPO Mix
prompts, intended for knowledge distillation into a student policy.
This is the combined dataset, merging:
big-math-ppo-mix-30k prompts → 22,584 correct solutions (75.3% of 30,000)
big-math-ppo-mix-extra-10k prompts → 7,986 correct solutions (79.9% of 10,000)
Total: 30,570 verified-correct solutions (40,000 prompts attempted).
Generation
vLLM sampling… See the full description on the dataset page: https://huggingface.co/datasets/christinakopi/big-math-ppo-mix-40k-qwq-solutions.
