CoolFace
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alexxxxxxia /Mako-DeepThink-RU Mako-DeepThink-RU (1.5B Tokens) Mako-DeepThink-RU — Russian CoT dataset. Because small models deserve to overthink too. Composition: 40% — Habr IT & Science (engineers explaining things) 30% — Russian CoT (step-by-step, no shortcuts) 20% — Wikipedia RU (curated, >1000 chars) 10% — Python-Edu (algorithms, clean) License License: Apache 2.0. Do whatever. Just don't blame us if your model starts reasoning at 3 AM. text1M<n<10M3 likes99 downloads7d agoHugging Face02open-llm-leaderboard /EpistemeAI__DeepThinkers-Phi4-detailsgated Dataset Card for Evaluation run of EpistemeAI/DeepThinkers-Phi4 Dataset automatically created during the evaluation run of model EpistemeAI/DeepThinkers-Phi4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__DeepThinkers-Phi4-details.tabular10K<n<100K0 likes73 downloads2y agoHugging Face03prithivMLmods /Deepthink-Reasoning Deepthink Reasoning Demo Deepthink Reasoning is a comprehensive data repository designed to break down complex problems, especially in coding (Python, Go, Java, C++, C#, etc.) and algorithms. It provides detailed problem analyses and systematic solutions to achieve the desired outcomes. Features Comprehensive Problem Breakdown: Deepthink Reasoning dissects problems into smaller, manageable components to facilitate effective understanding and solution generation.… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning.texttext-generationn<1K26 likes71 downloads1y agoHugging Face04open-llm-leaderboard /bunnycore__DeepThinker-7B-Sce-v1-detailsgated Dataset Card for Evaluation run of bunnycore/DeepThinker-7B-Sce-v1 Dataset automatically created during the evaluation run of model bunnycore/DeepThinker-7B-Sce-v1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__DeepThinker-7B-Sce-v1-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face05open-llm-leaderboard /DavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-detailsgated Dataset Card for Evaluation run of DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B Dataset automatically created during the evaluation run of model DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-details.tabular10K<n<100K2 likes41 downloads2y agoHugging Face06open-llm-leaderboard /prithivMLmods__Deepthink-Llama-3-8B-Preview-detailsgated Dataset Card for Evaluation run of prithivMLmods/Deepthink-Llama-3-8B-Preview Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Llama-3-8B-Preview The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Llama-3-8B-Preview-details.tabular10K<n<100K0 likes31 downloads2y agoHugging Face07prithivMLmods /Deepthink-Reasoning-Tamil Deepthink-Reasoning-Tamil Dataset The Deepthink-Reasoning-Tamil dataset is a multilingual dataset that includes both Tamil and Tanglish. It is designed with a dynamic instruction set, making it adaptable for various reasoning-based problem-solving tasks. Key Features: Multilingual Support: Includes Tamil and Tanglish for broader accessibility. Dynamic Instructions: Designed to adapt to various problem-solving tasks. Advanced Translation Pipeline: The dataset was… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning-Tamil.texttext-generationn<1K4 likes27 downloads2y agoHugging Face08Daemontatox /Deepthinking-COTtextn<1K1 likes26 downloads2y agoHugging Face09prithivMLmods /Deepthink-Reasoning-Instruction Deepthink Reasoning Demo Deepthink Reasoning is a comprehensive data repository designed to break down complex problems, especially in coding (Python, Go, Java, C++, C#, etc.) and algorithms. It provides detailed problem analyses and systematic solutions to achieve the desired outcomes. Features Comprehensive Problem Breakdown: Deepthink Reasoning dissects problems into smaller, manageable components to facilitate effective understanding and solution generation.… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning-Instruction.texttext-generationn<1K3 likes21 downloads1y agoHugging Face10Jessylg27 /DeepThink-Code-Lite 🧠 DeepThink Code - Lite Version (Reasoning & Code) ⚠️ Note : Ceci est la version LITE (645 exemples) destinée à l'évaluation et à la recherche non-commerciale. 🚀 Pour la version complète (20k+ exemples) avec Licence Commerciale, Cliquez ici pour accéder à l'offre complète sur Gumroad Description Ce dataset est conçu pour entraîner des modèles de langage à raisonner avant de coder. Contrairement aux datasets classiques qui donnent juste la solution, celui-ci force… See the full description on the dataset page: https://huggingface.co/datasets/Jessylg27/DeepThink-Code-Lite.texttext-generationn<1K1 likes21 downloads8mo agoHugging Face11open-llm-leaderboard /prithivMLmods__Deepthink-Reasoning-7B-detailsgated Dataset Card for Evaluation run of prithivMLmods/Deepthink-Reasoning-7B Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Reasoning-7B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Reasoning-7B-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face12Mgmgrand420 /DeepThink-Code-Lite 🧠 DeepThink Code - Lite Version (Reasoning & Code) ⚠️ Note : Ceci est la version LITE (645 exemples) destinée à l'évaluation et à la recherche non-commerciale. 🚀 Pour la version complète (20k+ exemples) avec Licence Commerciale, Cliquez ici pour accéder à l'offre complète sur Gumroad Description Ce dataset est conçu pour entraîner des modèles de langage à raisonner avant de coder. Contrairement aux datasets classiques qui donnent juste la solution, celui-ci force… See the full description on the dataset page: https://huggingface.co/datasets/Mgmgrand420/DeepThink-Code-Lite.texttext-generationn<1K0 likes19 downloads8mo agoHugging Face13open-llm-leaderboard /bunnycore__DeepThinker-7B-Sce-v2-detailsgated Dataset Card for Evaluation run of bunnycore/DeepThinker-7B-Sce-v2 Dataset automatically created during the evaluation run of model bunnycore/DeepThinker-7B-Sce-v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__DeepThinker-7B-Sce-v2-details.tabular10K<n<100K0 likes17 downloads2y agoHugging Face14mark-22 /Deepthinking-alfworld_and_dbbench_spider_v2 Deepthinking ALFWorld & DBBench Spider v2 AgentBench 評価の 2 タスク(ALFWorld / DBBench)を統合した マルチタスク SFT 訓練データセット。 フォーマット検査・フィルタリング済みの 7,779 件を、サイズ比率に基づく等間隔インターリーブで結合。 Dataset Summary Metric Value Total rows 7,779 ALFWorld 4,884 (62.8%) DBBench 2,895 (37.2%) Avg messages per item 18.3 Columns messages Interleave method 比率ベース等間隔マージ Source Datasets Source Rows Description mark-22/Deepthinking-sft_alfworld_final1 4,884 ALFWorld… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/Deepthinking-alfworld_and_dbbench_spider_v2.texttext-generation1K<n<10K0 likes16 downloads7mo agoHugging Face15mark-22 /Deepthinking-sft_alfworld_final1 Deepthinking SFT ALFWorld (Final) mark-22/alfworld_combined_shuffled_final(4,884 件)に対し、 GPT-OSS-120B (Groq) を用いて 最初の THOUGHT に「状況分析・常識推論・ステップ分解」を自動挿入 したデータ拡張版。 エージェントが行動前に深く考える(Deep Thinking)能力を SFT で獲得させることを目的としている。 Dataset Summary Metric Value Total rows 4,884 Source mark-22/alfworld_combined_shuffled_final Augmentation model GPT-OSS-120B (via Groq API) Avg messages per item 23.0 Items with THOUGHT + ACTION 4,884 / 4,884 (100%) Columns original_id… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/Deepthinking-sft_alfworld_final1.texttext-generation1K<n<10K0 likes13 downloads7mo agoHugging Face16open-llm-leaderboard /prithivMLmods__Deepthink-Reasoning-14B-detailsgated Dataset Card for Evaluation run of prithivMLmods/Deepthink-Reasoning-14B Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Reasoning-14B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Reasoning-14B-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face17mark-22 /Deepthinking-alfworld_and_dbbench_spider_v1text1K<n<10K0 likes10 downloads7mo agoHugging Face18future7 /deep_math_plus_8_with_retrieval_thinktext10K<n<100K0 likes9 downloads1y agoHugging Face19mark-22 /Deepthinking-sft_alfworld_v4text1K<n<10K0 likes7 downloads7mo agoHugging Face20open-llm-leaderboard /bunnycore__Llama-3.2-3B-RP-DeepThink-detailsgated Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-RP-DeepThink Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-RP-DeepThink The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-RP-DeepThink-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face21deepthink8 /pexelsphotosjanpfimage0 likes6 downloads11mo agoHugging Face22mark-22 /Deepthinking-sft_alfworld_test2textn<1K0 likes4 downloads7mo agoHugging Face23minyichen /Deepthink_R1gatedtabularn<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.