CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ChinaunicomSoftware /smoltalk-chinese-QwQ-Distrill smoltalk-chinese-QwQ-Distrill [中文] [English] 📖Technical Report smoltalk-chinese-QwQ-Distrill is a Chinese fine-tuning dataset constructed with reference to the SmolTalk-Chinese dataset. It aims to provide high-quality synthetic reasoning data support for training large language models (LLMs). The dataset consists entirely of synthetic data, comprising over 700,000 entries. It is specifically designed to enhance the performance of Chinese LLMs across various tasks… See the full description on the dataset page: https://huggingface.co/datasets/ChinaunicomSoftware/smoltalk-chinese-QwQ-Distrill.tabulartext-generation100K<n<1M3 likes180 downloads2y agoHugging Face02Tiiny /QWQ-LONGCOT-500KThis repository contains approximately 500,000 instances of responses generated using QwQ-32B-Preview language model. The dataset combines prompts from multiple high-quality sources to create diverse and comprehensive training data. The dataset is available under the Apache 2.0 license. Over 75% of the responses exceed 8,000 tokens in length. The majority of prompts were carefully created using persona-based methods to create challenging instructions. Bias, Risks, and Limitations… See the full description on the dataset page: https://huggingface.co/datasets/Tiiny/QWQ-LONGCOT-500K.text100K<n<1M124 likes121 downloads2y agoHugging Face03qwqeqw /Dataset_of_Russian_thinkingRu RTD Описание:Russian Thinking Dataset — это набор данных, предназначенный для обучения и тестирования моделей обработки естественного языка (NLP) на русском языке. Датасет ориентирован на задачи, связанные с генерацией текста, анализом диалогов и решением математических и логических задач. Основная информация: Сплит: train Количество записей: 147.046 Цели: Обучение моделей пониманию русского языка. Создание диалоговых систем с естественным взаимодействием.… See the full description on the dataset page: https://huggingface.co/datasets/qwqeqw/Dataset_of_Russian_thinking.texttext-generation100K<n<1M1 likes84 downloads10mo agoHugging Face04open-llm-leaderboard /Qwen__QwQ-32B-detailsgated Dataset Card for Evaluation run of Qwen/QwQ-32B Dataset automatically created during the evaluation run of model Qwen/QwQ-32B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-details.tabular10K<n<100K0 likes61 downloads2y agoHugging Face05open-llm-leaderboard /Qwen__QwQ-32B-Preview-detailsgated Dataset Card for Evaluation run of Qwen/QwQ-32B-Preview Dataset automatically created during the evaluation run of model Qwen/QwQ-32B-Preview The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Qwen__QwQ-32B-Preview-details.tabular10K<n<100K0 likes58 downloads2y agoHugging Face06WbjuSrceu /QwQ_32B_Preview_language_model_generation_not_confirmedtext100K<n<1M0 likes58 downloads2y agoHugging Face07open-llm-leaderboard /benhaotang__phi4-qwq-sky-t1-detailsgated Dataset Card for Evaluation run of benhaotang/phi4-qwq-sky-t1 Dataset automatically created during the evaluation run of model benhaotang/phi4-qwq-sky-t1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/benhaotang__phi4-qwq-sky-t1-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face08DeL-TaiseiOzaki /Tengentoppa-sft-seed-elyza-reasoning-QwQtext10K<n<100K0 likes46 downloads2y agoHugging Face09open-llm-leaderboard /FINGU-AI__QwQ-Buddy-32B-Alpha-detailsgated Dataset Card for Evaluation run of FINGU-AI/QwQ-Buddy-32B-Alpha Dataset automatically created during the evaluation run of model FINGU-AI/QwQ-Buddy-32B-Alpha The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FINGU-AI__QwQ-Buddy-32B-Alpha-details.tabular10K<n<100K0 likes46 downloads2y agoHugging Face10open-llm-leaderboard /bunnycore__QwQen-3B-LCoT-R1-detailsgated Dataset Card for Evaluation run of bunnycore/QwQen-3B-LCoT-R1 Dataset automatically created during the evaluation run of model bunnycore/QwQen-3B-LCoT-R1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__QwQen-3B-LCoT-R1-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face11open-llm-leaderboard /OpenBuddy__openbuddy-qwq-32b-v24.2-200k-detailsgated Dataset Card for Evaluation run of OpenBuddy/openbuddy-qwq-32b-v24.2-200k Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-qwq-32b-v24.2-200k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/OpenBuddy__openbuddy-qwq-32b-v24.2-200k-details.tabular10K<n<100K0 likes43 downloads2y agoHugging Face12open-llm-leaderboard /Daemontatox__Mini_QwQ-detailsgated Dataset Card for Evaluation run of Daemontatox/Mini_QwQ Dataset automatically created during the evaluation run of model Daemontatox/Mini_QwQ The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__Mini_QwQ-details.tabular10K<n<100K0 likes38 downloads2y agoHugging Face13xl-zhao /PromptCoT-QwQ-Dataset Dataset Format Each row in the dataset contains: prompt: The input to the reasoning model, including a problem statement with the special prompting template. completion: The expected output for supervised fine-tuning, containing a thought process wrapped in <think>...</think>, followed by the final solution. Example { "prompt": "<|im_start|>user\nLet $P$ be a point on a regular $n$-gon. A marker is placed on the vertex $P$ and a random process is used to… See the full description on the dataset page: https://huggingface.co/datasets/xl-zhao/PromptCoT-QwQ-Dataset.text10K<n<100K6 likes33 downloads1y agoHugging Face14QwQ2 /RoboPIN-Datasetstext100K<n<1M0 likes29 downloads3mo agoHugging Face15niryuu /Magpie-QwQ-Reasoning-dpoThis DPO dataset is synthesized by this way: instructions: Magpie on QwQ-32b-Preview outputs: QwQ-32b-Preview and calm3-22b-chat. textn<1K0 likes27 downloads2y agoHugging Face16demystify-long-cot /math-train-qwq-rs-n256text1M<n<10M1 likes25 downloads2y agoHugging Face17demystify-long-cot /math-train-qwq-rs-n192text100K<n<1M0 likes24 downloads2y agoHugging Face18rl-rag /qwq_32b_factualqa_sft_datatabular10K<n<100K0 likes21 downloads1y agoHugging Face19DataShare /QwQ_32B_setting_7text1K<n<10K0 likes21 downloads1y agoHugging Face20tugstugi /OpenMathInstruct-2-QwQ OpenMathInstruct-2-QwQ Qwen/QwQ-32B-Preview solutions of 107k augmented_math problems of nvidia/OpenMathInstruct-2. All solutions are validated and agree with the expected_answer field of OpenMathInstruct-2. texttext-generation100K<n<1M0 likes19 downloads2y agoHugging Face21huihui-ai /QWQ-LONGCOT-500KThis dataset is a copy of PowerInfer/QWQ-LONGCOT-500K. This repository contains approximately 500,000 instances of responses generated using QwQ-32B-Preview language model. The dataset combines prompts from multiple high-quality sources to create diverse and comprehensive training data. The dataset is available under the Apache 2.0 license. Over 75% of the responses exceed 8,000 tokens in length. The majority of prompts were carefully created using persona-based methods to create challenging… See the full description on the dataset page: https://huggingface.co/datasets/huihui-ai/QWQ-LONGCOT-500K.text100K<n<1M2 likes16 downloads2y agoHugging Face22DeL-TaiseiOzaki /Tengentoppa-QwQ-reasoning-sft-elyzatext10K<n<100K0 likes15 downloads2y agoHugging Face23PSM24 /qwq-MATH500-94textn<1K0 likes14 downloads2y agoHugging Face24dmitriihook /blocksworld-mystery-4-qwq-reasoning-parts-explorationtextn<1K0 likes14 downloads2y agoHugging Face25au6000 /QwQ32B_GAIR_LIMO_v2text1K<n<10K0 likes14 downloads1y agoHugging Face26open-llm-leaderboard /prithivMLmods__QwQ-LCoT-14B-Conversational-detailsgated Dataset Card for Evaluation run of prithivMLmods/QwQ-LCoT-14B-Conversational Dataset automatically created during the evaluation run of model prithivMLmods/QwQ-LCoT-14B-Conversational The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__QwQ-LCoT-14B-Conversational-details.tabular10K<n<100K1 likes13 downloads2y agoHugging Face27demystify-long-cot /math-train-qwq-rs-n128text100K<n<1M1 likes13 downloads2y agoHugging Face28open-llm-leaderboard /qingy2024__QwQ-14B-Math-v0.2-detailsgated Dataset Card for Evaluation run of qingy2024/QwQ-14B-Math-v0.2 Dataset automatically created during the evaluation run of model qingy2024/QwQ-14B-Math-v0.2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__QwQ-14B-Math-v0.2-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face29open-llm-leaderboard /Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-detailsgated Dataset Card for Evaluation run of Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ Dataset automatically created during the evaluation run of model Pinkstack/SuperThoughts-CoT-14B-16k-o1-QwQ The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Pinkstack__SuperThoughts-CoT-14B-16k-o1-QwQ-details.tabular10K<n<100K2 likes12 downloads2y agoHugging Face30yinjiewang /prun-QwQ-32B-MATH_traintext1K<n<10K0 likes12 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.