CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ChinaunicomSoftware /smoltalk-chinese-QwQ-Distrill smoltalk-chinese-QwQ-Distrill [中文] [English] 📖Technical Report smoltalk-chinese-QwQ-Distrill is a Chinese fine-tuning dataset constructed with reference to the SmolTalk-Chinese dataset. It aims to provide high-quality synthetic reasoning data support for training large language models (LLMs). The dataset consists entirely of synthetic data, comprising over 700,000 entries. It is specifically designed to enhance the performance of Chinese LLMs across various tasks… See the full description on the dataset page: https://huggingface.co/datasets/ChinaunicomSoftware/smoltalk-chinese-QwQ-Distrill.tabulartext-generation100K<n<1M3 likes159 downloads2y agoHugging Face02Tiiny /QWQ-LONGCOT-500KThis repository contains approximately 500,000 instances of responses generated using QwQ-32B-Preview language model. The dataset combines prompts from multiple high-quality sources to create diverse and comprehensive training data. The dataset is available under the Apache 2.0 license. Over 75% of the responses exceed 8,000 tokens in length. The majority of prompts were carefully created using persona-based methods to create challenging instructions. Bias, Risks, and Limitations… See the full description on the dataset page: https://huggingface.co/datasets/Tiiny/QWQ-LONGCOT-500K.text100K<n<1M124 likes120 downloads2y agoHugging Face03qwqeqw /Dataset_of_Russian_thinkingRu RTD Описание:Russian Thinking Dataset — это набор данных, предназначенный для обучения и тестирования моделей обработки естественного языка (NLP) на русском языке. Датасет ориентирован на задачи, связанные с генерацией текста, анализом диалогов и решением математических и логических задач. Основная информация: Сплит: train Количество записей: 147.046 Цели: Обучение моделей пониманию русского языка. Создание диалоговых систем с естественным взаимодействием.… See the full description on the dataset page: https://huggingface.co/datasets/qwqeqw/Dataset_of_Russian_thinking.texttext-generation100K<n<1M1 likes86 downloads10mo agoHugging Face04open-llm-leaderboard /Qwen__QwQ-32B-detailsgated Dataset Card for Evaluation run of Qwen/QwQ-32B Dataset automatically created during the evaluation run of model Qwen/QwQ-32B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-details.tabular10K<n<100K0 likes69 downloads2y agoHugging Face05WbjuSrceu /QwQ_32B_Preview_language_model_generation_not_confirmedtext100K<n<1M0 likes57 downloads2y agoHugging Face06open-llm-leaderboard /Qwen__QwQ-32B-Preview-detailsgated Dataset Card for Evaluation run of Qwen/QwQ-32B-Preview Dataset automatically created during the evaluation run of model Qwen/QwQ-32B-Preview The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-Preview-details.tabular10K<n<100K0 likes56 downloads2y agoHugging Face07open-llm-leaderboard /benhaotang__phi4-qwq-sky-t1-detailsgated Dataset Card for Evaluation run of benhaotang/phi4-qwq-sky-t1 Dataset automatically created during the evaluation run of model benhaotang/phi4-qwq-sky-t1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/benhaotang__phi4-qwq-sky-t1-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face08open-llm-leaderboard /FINGU-AI__QwQ-Buddy-32B-Alpha-detailsgated Dataset Card for Evaluation run of FINGU-AI/QwQ-Buddy-32B-Alpha Dataset automatically created during the evaluation run of model FINGU-AI/QwQ-Buddy-32B-Alpha The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FINGU-AI__QwQ-Buddy-32B-Alpha-details.tabular10K<n<100K0 likes46 downloads2y agoHugging Face09open-llm-leaderboard /bunnycore__QwQen-3B-LCoT-R1-detailsgated Dataset Card for Evaluation run of bunnycore/QwQen-3B-LCoT-R1 Dataset automatically created during the evaluation run of model bunnycore/QwQen-3B-LCoT-R1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__QwQen-3B-LCoT-R1-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face10open-llm-leaderboard /OpenBuddy__openbuddy-qwq-32b-v24.2-200k-detailsgated Dataset Card for Evaluation run of OpenBuddy/openbuddy-qwq-32b-v24.2-200k Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-qwq-32b-v24.2-200k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/OpenBuddy__openbuddy-qwq-32b-v24.2-200k-details.tabular10K<n<100K0 likes43 downloads2y agoHugging Face11DeL-TaiseiOzaki /Tengentoppa-sft-seed-elyza-reasoning-QwQtext10K<n<100K0 likes37 downloads2y agoHugging Face12open-llm-leaderboard /Daemontatox__Mini_QwQ-detailsgated Dataset Card for Evaluation run of Daemontatox/Mini_QwQ Dataset automatically created during the evaluation run of model Daemontatox/Mini_QwQ The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__Mini_QwQ-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face13QwQ2 /RoboPIN-Datasetstext100K<n<1M0 likes34 downloads3mo agoHugging Face14xl-zhao /PromptCoT-QwQ-Dataset Dataset Format Each row in the dataset contains: prompt: The input to the reasoning model, including a problem statement with the special prompting template. completion: The expected output for supervised fine-tuning, containing a thought process wrapped in <think>...</think>, followed by the final solution. Example { "prompt": "<|im_start|>user\nLet $P$ be a point on a regular $n$-gon. A marker is placed on the vertex $P$ and a random process is used to… See the full description on the dataset page: https://huggingface.co/datasets/xl-zhao/PromptCoT-QwQ-Dataset.text10K<n<100K6 likes33 downloads1y agoHugging Face15niryuu /Magpie-QwQ-Reasoning-dpoThis DPO dataset is synthesized by this way: instructions: Magpie on QwQ-32b-Preview outputs: QwQ-32b-Preview and calm3-22b-chat. textn<1K0 likes26 downloads2y agoHugging Face16demystify-long-cot /math-train-qwq-rs-n192text100K<n<1M0 likes25 downloads2y agoHugging Face17rl-rag /qwq_32b_factualqa_sft_datatabular10K<n<100K0 likes24 downloads1y agoHugging Face18demystify-long-cot /math-train-qwq-rs-n256text1M<n<10M1 likes22 downloads2y agoHugging Face19DataShare /QwQ_32B_setting_7text1K<n<10K0 likes21 downloads1y agoHugging Face20tugstugi /OpenMathInstruct-2-QwQ OpenMathInstruct-2-QwQ Qwen/QwQ-32B-Preview solutions of 107k augmented_math problems of nvidia/OpenMathInstruct-2. All solutions are validated and agree with the expected_answer field of OpenMathInstruct-2. texttext-generation100K<n<1M0 likes19 downloads2y agoHugging Face21huihui-ai /QWQ-LONGCOT-500KThis dataset is a copy of PowerInfer/QWQ-LONGCOT-500K. This repository contains approximately 500,000 instances of responses generated using QwQ-32B-Preview language model. The dataset combines prompts from multiple high-quality sources to create diverse and comprehensive training data. The dataset is available under the Apache 2.0 license. Over 75% of the responses exceed 8,000 tokens in length. The majority of prompts were carefully created using persona-based methods to create challenging… See the full description on the dataset page: https://huggingface.co/datasets/huihui-ai/QWQ-LONGCOT-500K.text100K<n<1M2 likes15 downloads2y agoHugging Face22PSM24 /qwq-MATH500-94textn<1K0 likes14 downloads2y agoHugging Face23au6000 /QwQ32B_GAIR_LIMO_v2text1K<n<10K0 likes14 downloads1y agoHugging Face24open-llm-leaderboard /Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-detailsgated Dataset Card for Evaluation run of Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ Dataset automatically created during the evaluation run of model Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-details.tabular10K<n<100K2 likes13 downloads2y agoHugging Face25open-llm-leaderboard /prithivMLmods__QwQ-LCoT-14B-Conversational-detailsgated Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT-14B-Conversational Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT-14B-Conversational The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT-14B-Conversational-details.tabular10K<n<100K1 likes13 downloads2y agoHugging Face26demystify-long-cot /math-train-qwq-rs-n128text100K<n<1M1 likes13 downloads2y agoHugging Face27open-llm-leaderboard /prithivMLmods__QwQ-R1-Distill-1.5B-CoT-detailsgated Dataset Card for Evaluation run of prithivMLmods/QwQ-R1-Distill-1.5B-CoT Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-R1-Distill-1.5B-CoT The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-R1-Distill-1.5B-CoT-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face28VanWang /NuminaMath-CoT_O1_Qwqtext1K<n<10K0 likes12 downloads2y agoHugging Face29dmitriihook /blocksworld-mystery-4-qwq-reasoning-parts-explorationtextn<1K0 likes12 downloads2y agoHugging Face30yinjiewang /prun-QwQ-32B-MATH_traintext1K<n<10K0 likes12 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.