datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
kling-ai-imagesMako-DeepThink-RU
Mako-DeepThink-RU (1.5B Tokens)
Mako-DeepThink-RU — Russian CoT dataset. Because small models deserve to overthink too.
Composition:
40% — Habr IT & Science (engineers explaining things)
30% — Russian CoT (step-by-step, no shortcuts)
20% — Wikipedia RU (curated, >1000 chars)
10% — Python-Edu (algorithms, clean)
License
License: Apache 2.0. Do whatever. Just don't blame us if your model starts reasoning at 3 AM.
EpistemeAI__DeepThinkers-Phi4-details
Dataset Card for Evaluation run of EpistemeAI/DeepThinkers-Phi4
Dataset automatically created during the evaluation run of model EpistemeAI/DeepThinkers-Phi4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__DeepThinkers-Phi4-details.Deepthink-Reasoning
Deepthink Reasoning Demo
Deepthink Reasoning is a comprehensive data repository designed to break down complex problems, especially in coding (Python, Go, Java, C++, C#, etc.) and algorithms. It provides detailed problem analyses and systematic solutions to achieve the desired outcomes.
Features
Comprehensive Problem Breakdown: Deepthink Reasoning dissects problems into smaller, manageable components to facilitate effective understanding and solution generation.… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning.deep-thinkbunnycore__DeepThinker-7B-Sce-v1-details
Dataset Card for Evaluation run of bunnycore/DeepThinker-7B-Sce-v1
Dataset automatically created during the evaluation run of model bunnycore/DeepThinker-7B-Sce-v1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__DeepThinker-7B-Sce-v1-details.sample-real-fake-data-ImageDavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-details
Dataset Card for Evaluation run of DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B
Dataset automatically created during the evaluation run of model DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-details.details_xDAN-AI__xDAN-L1Mix-DeepThinking-v2
Dataset Card for Evaluation run of xDAN-AI/xDAN-L1Mix-DeepThinking-v2
Dataset automatically created during the evaluation run of model xDAN-AI/xDAN-L1Mix-DeepThinking-v2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_xDAN-AI__xDAN-L1Mix-DeepThinking-v2.Whab-deepfake-imageprithivMLmods__Deepthink-Llama-3-8B-Preview-details
Dataset Card for Evaluation run of prithivMLmods/Deepthink-Llama-3-8B-Preview
Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Llama-3-8B-Preview
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Llama-3-8B-Preview-details.Deepthink-Reasoning-Tamil
Deepthink-Reasoning-Tamil Dataset
The Deepthink-Reasoning-Tamil dataset is a multilingual dataset that includes both Tamil and Tanglish. It is designed with a dynamic instruction set, making it adaptable for various reasoning-based problem-solving tasks.
Key Features:
Multilingual Support: Includes Tamil and Tanglish for broader accessibility.
Dynamic Instructions: Designed to adapt to various problem-solving tasks.
Advanced Translation Pipeline: The dataset was… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning-Tamil.Deepthinking-COTDeepthink-Reasoning-Instruction
Deepthink Reasoning Demo
Deepthink Reasoning is a comprehensive data repository designed to break down complex problems, especially in coding (Python, Go, Java, C++, C#, etc.) and algorithms. It provides detailed problem analyses and systematic solutions to achieve the desired outcomes.
Features
Comprehensive Problem Breakdown: Deepthink Reasoning dissects problems into smaller, manageable components to facilitate effective understanding and solution generation.… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Deepthink-Reasoning-Instruction.Deepfake-vs-Real-60KDeepThink-Code-Lite
🧠 DeepThink Code - Lite Version (Reasoning & Code)
⚠️ Note : Ceci est la version LITE (645 exemples) destinée à l'évaluation et à la recherche non-commerciale.
🚀 Pour la version complète (20k+ exemples) avec Licence Commerciale, Cliquez ici pour accéder à l'offre complète sur Gumroad
Description
Ce dataset est conçu pour entraîner des modèles de langage à raisonner avant de coder. Contrairement aux datasets classiques qui donnent juste la solution, celui-ci force… See the full description on the dataset page: https://huggingface.co/datasets/Jessylg27/DeepThink-Code-Lite.prithivMLmods__Deepthink-Reasoning-7B-details
Dataset Card for Evaluation run of prithivMLmods/Deepthink-Reasoning-7B
Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Reasoning-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Reasoning-7B-details.Flickr-Faces-HQ-DatasetDeepThink-Code-Lite
🧠 DeepThink Code - Lite Version (Reasoning & Code)
⚠️ Note : Ceci est la version LITE (645 exemples) destinée à l'évaluation et à la recherche non-commerciale.
🚀 Pour la version complète (20k+ exemples) avec Licence Commerciale, Cliquez ici pour accéder à l'offre complète sur Gumroad
Description
Ce dataset est conçu pour entraîner des modèles de langage à raisonner avant de coder. Contrairement aux datasets classiques qui donnent juste la solution, celui-ci force… See the full description on the dataset page: https://huggingface.co/datasets/Mgmgrand420/DeepThink-Code-Lite.bunnycore__DeepThinker-7B-Sce-v2-details
Dataset Card for Evaluation run of bunnycore/DeepThinker-7B-Sce-v2
Dataset automatically created during the evaluation run of model bunnycore/DeepThinker-7B-Sce-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__DeepThinker-7B-Sce-v2-details.Deepthinking-alfworld_and_dbbench_spider_v2
Deepthinking ALFWorld & DBBench Spider v2
AgentBench 評価の 2 タスク(ALFWorld / DBBench)を統合した マルチタスク SFT 訓練データセット。
フォーマット検査・フィルタリング済みの 7,779 件を、サイズ比率に基づく等間隔インターリーブで結合。
Dataset Summary
Metric
Value
Total rows
7,779
ALFWorld
4,884 (62.8%)
DBBench
2,895 (37.2%)
Avg messages per item
18.3
Columns
messages
Interleave method
比率ベース等間隔マージ
Source Datasets
Source
Rows
Description
mark-22/Deepthinking-sft_alfworld_final1
4,884
ALFWorld… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/Deepthinking-alfworld_and_dbbench_spider_v2.Deepthinking-sft_alfworld_final1
Deepthinking SFT ALFWorld (Final)
mark-22/alfworld_combined_shuffled_final(4,884 件)に対し、
GPT-OSS-120B (Groq) を用いて 最初の THOUGHT に「状況分析・常識推論・ステップ分解」を自動挿入 したデータ拡張版。
エージェントが行動前に深く考える(Deep Thinking)能力を SFT で獲得させることを目的としている。
Dataset Summary
Metric
Value
Total rows
4,884
Source
mark-22/alfworld_combined_shuffled_final
Augmentation model
GPT-OSS-120B (via Groq API)
Avg messages per item
23.0
Items with THOUGHT + ACTION
4,884 / 4,884 (100%)
Columns
original_id… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/Deepthinking-sft_alfworld_final1.ShareGPT4VideoprithivMLmods__Deepthink-Reasoning-14B-details
Dataset Card for Evaluation run of prithivMLmods/Deepthink-Reasoning-14B
Dataset automatically created during the evaluation run of model prithivMLmods/Deepthink-Reasoning-14B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Deepthink-Reasoning-14B-details.receipts_even_moreOnePromptOneStory-RunDiffusion-Juggernaut-X-v10Deepthinking-alfworld_and_dbbench_spider_v1deep_math_plus_8_with_retrieval_thinkreceipt-koImgreceipts-finetune-v3
