datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
smoltalk-chinese-QwQ-Distrill
smoltalk-chinese-QwQ-Distrill [中文] [English]
📖Technical Report
smoltalk-chinese-QwQ-Distrill is a Chinese fine-tuning dataset constructed with reference to the SmolTalk-Chinese dataset. It aims to provide high-quality synthetic reasoning data support for training large language models (LLMs). The dataset consists entirely of synthetic data, comprising over 700,000 entries. It is specifically designed to enhance the performance of Chinese LLMs across various tasks… See the full description on the dataset page: https://huggingface.co/datasets/ChinaunicomSoftware/smoltalk-chinese-QwQ-Distrill.Qwen__QwQ-32B-details
Dataset Card for Evaluation run of Qwen/QwQ-32B
Dataset automatically created during the evaluation run of model Qwen/QwQ-32B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-details.Qwen__QwQ-32B-Preview-details
Dataset Card for Evaluation run of Qwen/QwQ-32B-Preview
Dataset automatically created during the evaluation run of model Qwen/QwQ-32B-Preview
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-Preview-details.benhaotang__phi4-qwq-sky-t1-details
Dataset Card for Evaluation run of benhaotang/phi4-qwq-sky-t1
Dataset automatically created during the evaluation run of model benhaotang/phi4-qwq-sky-t1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/benhaotang__phi4-qwq-sky-t1-details.FINGU-AI__QwQ-Buddy-32B-Alpha-details
Dataset Card for Evaluation run of FINGU-AI/QwQ-Buddy-32B-Alpha
Dataset automatically created during the evaluation run of model FINGU-AI/QwQ-Buddy-32B-Alpha
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FINGU-AI__QwQ-Buddy-32B-Alpha-details.bunnycore__QwQen-3B-LCoT-R1-details
Dataset Card for Evaluation run of bunnycore/QwQen-3B-LCoT-R1
Dataset automatically created during the evaluation run of model bunnycore/QwQen-3B-LCoT-R1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__QwQen-3B-LCoT-R1-details.OpenBuddy__openbuddy-qwq-32b-v24.2-200k-details
Dataset Card for Evaluation run of OpenBuddy/openbuddy-qwq-32b-v24.2-200k
Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-qwq-32b-v24.2-200k
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/OpenBuddy__openbuddy-qwq-32b-v24.2-200k-details.Daemontatox__Mini_QwQ-details
Dataset Card for Evaluation run of Daemontatox/Mini_QwQ
Dataset automatically created during the evaluation run of model Daemontatox/Mini_QwQ
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__Mini_QwQ-details.qwq_32b_factualqa_sft_dataprithivMLmods__QwQ-LCoT-14B-Conversational-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT-14B-Conversational
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT-14B-Conversational
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT-14B-Conversational-details.qingy2024__QwQ-14B-Math-v0.2-details
Dataset Card for Evaluation run of qingy2024/QwQ-14B-Math-v0.2
Dataset automatically created during the evaluation run of model qingy2024/QwQ-14B-Math-v0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__QwQ-14B-Math-v0.2-details.Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-details
Dataset Card for Evaluation run of Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ
Dataset automatically created during the evaluation run of model Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-details.sdconfigprithivMLmods__QwQ-LCoT-7B-Instruct-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT-7B-Instruct
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT-7B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT-7B-Instruct-details.prithivMLmods__QwQ-R1-Distill-1.5B-CoT-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-R1-Distill-1.5B-CoT
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-R1-Distill-1.5B-CoT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-R1-Distill-1.5B-CoT-details.bunnycore__QwQen-3B-LCoT-details
Dataset Card for Evaluation run of bunnycore/QwQen-3B-LCoT
Dataset automatically created during the evaluation run of model bunnycore/QwQen-3B-LCoT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__QwQen-3B-LCoT-details.huihui-ai__QwQ-32B-Coder-Fusion-8020-details
Dataset Card for Evaluation run of huihui-ai/QwQ-32B-Coder-Fusion-8020
Dataset automatically created during the evaluation run of model huihui-ai/QwQ-32B-Coder-Fusion-8020
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/huihui-ai__QwQ-32B-Coder-Fusion-8020-details.huihui-ai__QwQ-32B-Coder-Fusion-7030-details
Dataset Card for Evaluation run of huihui-ai/QwQ-32B-Coder-Fusion-7030
Dataset automatically created during the evaluation run of model huihui-ai/QwQ-32B-Coder-Fusion-7030
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/huihui-ai__QwQ-32B-Coder-Fusion-7030-details.prithivMLmods__Phi-4-QwQ-details
Dataset Card for Evaluation run of prithivMLmods/Phi-4-QwQ
Dataset automatically created during the evaluation run of model prithivMLmods/Phi-4-QwQ
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Phi-4-QwQ-details.prithivMLmods__QwQ-MathOct-7B-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-MathOct-7B
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-MathOct-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-MathOct-7B-details.OpenBuddy__openbuddy-qwq-32b-v24.1-200k-details
Dataset Card for Evaluation run of OpenBuddy/openbuddy-qwq-32b-v24.1-200k
Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-qwq-32b-v24.1-200k
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/OpenBuddy__openbuddy-qwq-32b-v24.1-200k-details.tugstugi__Qwen2.5-7B-Instruct-QwQ-v0.1-details
Dataset Card for Evaluation run of tugstugi/Qwen2.5-7B-Instruct-QwQ-v0.1
Dataset automatically created during the evaluation run of model tugstugi/Qwen2.5-7B-Instruct-QwQ-v0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tugstugi__Qwen2.5-7B-Instruct-QwQ-v0.1-details.prithivMLmods__QwQ-LCoT2-7B-Instruct-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT2-7B-Instruct
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT2-7B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT2-7B-Instruct-details.QwQ32B_created_GAIR_LIMOv1kz919__QwQ-0.5B-Distilled-SFT-details
Dataset Card for Evaluation run of kz919/QwQ-0.5B-Distilled-SFT
Dataset automatically created during the evaluation run of model kz919/QwQ-0.5B-Distilled-SFT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kz919__QwQ-0.5B-Distilled-SFT-details.prithivMLmods__QwQ-LCoT-3B-Instruct-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT-3B-Instruct
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT-3B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT-3B-Instruct-details.Pinkstack__PARM-V1.5-base-QwQ-Qwen-2.5-o1-3B-details
Dataset Card for Evaluation run of Pinkstack/PARM-V1.5-base-QwQ-Qwen-2.5-o1-3B
Dataset automatically created during the evaluation run of model Pinkstack/PARM-V1.5-base-QwQ-Qwen-2.5-o1-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Pinkstack__PARM-V1.5-base-QwQ-Qwen-2.5-o1-3B-details.prithivMLmods__QwQ-R1-Distill-7B-CoT-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-R1-Distill-7B-CoT
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-R1-Distill-7B-CoT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-R1-Distill-7B-CoT-details.prithivMLmods__QwQ-LCoT1-Merged-details
Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT1-Merged
Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT1-Merged
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT1-Merged-details.sdconfig2
