CoolFace
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mlfoundations-dev /Nemotron-Research-Reasoning-Qwen-1.5B_eval_569atabular1K<n<10K0 likes1.1k downloads1y agoHugging Face02Kylan12 /mycotoxin-chemical-research-sythetic-reasoning mycotoxin-chemical-research-sythetic-reasoning Synthetic Q&A dataset on Mycotoxin Chemical Research, generated with SDGS (Synthetic Dataset Generation Suite). Dataset Details Metric Value Topic Mycotoxin Chemical Research Total Q&A Pairs 4416 Valid Pairs 4416 Provider/Model ollama/gpt-oss:120b Generation Cost Metric Value Prompt Tokens 4,579,253 Completion Tokens 5,326,284 Total Tokens 9,905,537 GPU Energy 3.5712 kWh… See the full description on the dataset page: https://huggingface.co/datasets/Kylan12/mycotoxin-chemical-research-sythetic-reasoning.textquestion-answering1K<n<10K0 likes198 downloads7mo agoHugging Face03ulamai /verified-research-reasoning-trajectories Verified Research Reasoning Trajectories for RLVR This repository is the public sample and schema repository for Ulam's research-level mathematical reasoning trajectories for reinforcement learning with verifiable rewards (RLVR), process supervision, judge training, proof criticism, and private evaluations. Ulam Verified Research Reasoning Trajectories are proof-process data for RLVR. Each record contains a normalized research problem, a golden or partial-golden proof graph… See the full description on the dataset page: https://huggingface.co/datasets/ulamai/verified-research-reasoning-trajectories.documenttext-generationn<1K3 likes174 downloads2mo agoHugging Face04AmanPriyanshu /tool-reasoning-sft-RESEARCH-grill-lab-browsecomp-plus-runs-data-cleaned-rectified Tool-Reasoning SFT — BrowseComp-Plus Runs (Cleaned & Rectified) Multi-turn tool-use reasoning trajectories derived from grill-lab/browsecomp-plus-runs, converted to a structured SFT format following the interstellarninja/hermes_reasoning_tool_use convention. Source Based on the execution trajectories from "Revisiting Text Ranking in Deep Research" (arXiv:2602.21456): Original data: grill-lab/browsecomp-plus-runs (MIT) Format Each row contains a messages… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-grill-lab-browsecomp-plus-runs-data-cleaned-rectified.texttext-generation10K<n<100K0 likes103 downloads7mo agoHugging Face05geodesic-research /persistent-alignment-warm-start-short-reasoning geodesic-research/persistent-alignment-warm-start-short-reasoning Local-pipeline snapshot published via --push-from-local. All configs below were built locally (Hub-independent) and uploaded in a single commit at one snapshot revision. Pipeline run params hash: 5350879c062dde0794a77181cebc05387828bff5efb326553c6c95eea675fa31 Configs in this snapshot: agentic_interactive, agentic_search, chat_multiturn, default, instruction_following, math_reasoning, safety, science_mcq… See the full description on the dataset page: https://huggingface.co/datasets/geodesic-research/persistent-alignment-warm-start-short-reasoning.tabular100K<n<1M0 likes92 downloads2mo agoHugging Face06SupritiVijay /tool-reasoning-sft-RESEARCH-dr-tulu-sft-deep-research-agent-data-cleaned-rectified Deep Research - Tulu SFT Data Cleaned Rectified 👥 Follow the Author Supriti Vijay Overview This dataset is a cleaned and restructured version of the DR-TULU SFT dataset released by AllenAI's RL Research team. The original DR-TULU dataset represents significant work in creating high-quality training data for reasoning-enhanced language models with tool use capabilities. This version addresses structural issues in the original release while preserving… See the full description on the dataset page: https://huggingface.co/datasets/SupritiVijay/tool-reasoning-sft-RESEARCH-dr-tulu-sft-deep-research-agent-data-cleaned-rectified.tabulartext-generation10K<n<100K8 likes87 downloads10mo agoHugging Face07AmanPriyanshu /tool-reasoning-sft-RESEARCH-OpenHands-CodeScout_Training_Rollouts CodeScout Training Rollouts — Cleaned & Rectified ~40K multi-turn code localization agent trajectories converted into a strict reasoning + tool-call format with validated FSM transitions. Supports coupled (parallel) tool calls. ⚠️ Mid-training dataset. This dataset contains synthesized reasoning templates (not native chain-of-thought). It is suitable for mid-training to teach tool-use mechanics, FSM structure, and bash exploration patterns. It is not recommended as a final SFT… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-OpenHands-CodeScout_Training_Rollouts.texttext-generation10K<n<100K0 likes74 downloads6mo agoHugging Face08AmanPriyanshu /tool-reasoning-sft-RESEARCH-rlvr-env-retrieval-source Tool-Reasoning SFT — RLVR Retrieval Source Trajectories 156,381 multi-turn agentic retrieval trajectories across three document corpora, in a strict reasoning + tool-call format with validated FSM transitions. Each trajectory records a model searching a corpus, opening documents, and citing relevant passages to answer a question. Author: Aman Priyanshu Source Environments Trajectories were collected against three RLVR retrieval environments from the FORMAT: Search -… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-rlvr-env-retrieval-source.tabulartext-generation100K<n<1M0 likes74 downloads6mo agoHugging Face09ciol-research /multilevel-legal-reasoning Legal Reasoning Dataset with Multilevel Human and Model-Annotated Explanations Prepared by Mst Rafia Islam, Umong Sain, Azmine Toushik Wasi Prepared as a part of Reasoning Datasets Competition by Bespoke Labs, Hugging Face, and Together.ai. 🧭 Purpose and Scope The Legal Reasoning Dataset aims to support the evaluation and training of legal reasoning systems, particularly in multilingual or jurisdiction-agnostic contexts. It focuses on international acts and treaties… See the full description on the dataset page: https://huggingface.co/datasets/ciol-research/multilevel-legal-reasoning.tabulartext-generationn<1K7 likes64 downloads1y agoHugging Face10AmanPriyanshu /tool-reasoning-sft-RESEARCH-explorations Explorations Trajectories — Cleaned & Stripped 149,025 multi-turn code exploration agent trajectories converted into a strict reasoning + tool-call format with validated FSM transitions. Origin Derived from AmanPriyanshu/random-small-github-repositories and AmanPriyanshu/random-python-github-repositories. Each trajectory is a search session where an agent navigates a GitHub repository using terminal commands to locate a target file. The agent reasons about project… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-explorations.tabulartext-generation100K<n<1M1 likes54 downloads6mo agoHugging Face11p-research /kimi-k3-cyber-reasoning-distill Kimi Cyber Reasoning 997 chain-of-thought records covering 13 cybersecurity disciplines and 4 systems engineering domains, distilled from the Kimi K3 reasoning model via API. Every record provides an explicit step-by-step <think> reasoning trace followed by a technical resolution, unified code diff fix, or structured tool invocation. The dataset was curated as an anchor set for training, healing, and specializing compact reasoning models on systems security and tool calling… See the full description on the dataset page: https://huggingface.co/datasets/p-research/kimi-k3-cyber-reasoning-distill.texttext-generationn<1K0 likes51 downloads10d agoHugging Face12Glint-Research /Opus-4.6-Reasoning-2160x Opus-4.6-Reasoning-2160x 2,160 high-quality reasoning traces generated by Claude Opus 4.6 via OpenRouter, covering mathematics, competitive programming, logic, science, and language tasks. Each example includes the full problem, an extended chain-of-thought, and a final solution — making the dataset suitable for supervised fine-tuning, chain-of-thought distillation, and reasoning-capability transfer to smaller models. Originally generated as a batch of 3,305 examples; 1,145 were… See the full description on the dataset page: https://huggingface.co/datasets/Glint-Research/Opus-4.6-Reasoning-2160x.texttext-generation1K<n<10K16 likes49 downloads5mo agoHugging Face13AmanPriyanshu /tool-reasoning-sft-RESEARCH-REDSearcher_SFT_10K REDSearcher SFT 10K — Cleaned & Rectified ~8,850 multi-turn deep-search agent trajectories converted into a strict reasoning + tool-call format with validated FSM transitions. Origin Derived from Zchu/REDSearcher_SFT_10K. REDSearcher is a deep search assistant dataset featuring rigorous, multi-step, multi-source investigations. Each trajectory contains a complex research question answered through iterative search → visit → reason → answer cycles, with extensive… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-REDSearcher_SFT_10K.texttext-generation1K<n<10K0 likes36 downloads6mo agoHugging Face14AmanPriyanshu /tool-reasoning-sft-RESEARCH-OpenSeeker-v1-Data OpenSeeker v1 — Cleaned & Rectified 7,189 multi-turn deep-search agent trajectories converted into a strict reasoning + tool-call format with validated FSM transitions. Origin Derived from OpenSeeker/OpenSeeker-v1-Data. OpenSeeker is an open-source search agent system that democratizes access to frontier search capabilities by fully open-sourcing its training data. Fine-tuned on Qwen3-30B-A3B-Thinking-2507 with 11.7K training examples, it achieves state-of-the-art… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-OpenSeeker-v1-Data.texttext-generation1K<n<10K0 likes36 downloads6mo agoHugging Face15aakashmallik /research-paper-agent-reasoning-traces-unverifiedtextn<1K0 likes30 downloads2mo agoHugging Face16geodesic-research /sfm-cpt-reasoning-comparetext10K<n<100K0 likes28 downloads7mo agoHugging Face17geodesic-research /sfm-cpt-reasoning-compare-pairedtext1K<n<10K0 likes22 downloads7mo agoHugging Face18orion-research /physical-reasoning-pt_BRtextn<1K0 likes15 downloads2y agoHugging Face19orion-research /general-reasoning-pt_BRtext1K<n<10K0 likes12 downloads2y agoHugging Face20mlfoundations-dev /Nemotron-Research-Reasoning-Qwen-1.5B_eval_5554 mlfoundations-dev/Nemotron-Research-Reasoning-Qwen-1.5B_eval_5554 Precomputed model outputs for evaluation. Evaluation Results Summary Metric AIME24 AMC23 MATH500 MMLUPro JEEBench GPQADiamond LiveCodeBench CodeElo CodeForces HLE HMMT AIME25 LiveCodeBenchv5 Accuracy 47.7 87.5 86.0 32.3 52.6 41.8 31.4 54.7 40.3 10.5 21.7 32.0 22.6 AIME24 Average Accuracy: 47.67% ± 1.49% Number of Runs: 10 Run Accuracy Questions Solved Total… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/Nemotron-Research-Reasoning-Qwen-1.5B_eval_5554.tabular10K<n<100K0 likes12 downloads1y agoHugging Face21LLMTeamAkiyama /cleand_moremilk_CoT_Reasoning_Scientific_Discovery_and_Research元データ: https://huggingface.co/datasets/moremilk/CoT_Reasoning_Scientific_Discovery_and_Research 使用したコード: https://github.com/LLMTeamAkiyama/0-data_prepare/tree/master/src/CoT_Reasoning_Scientific_Discovery_and_Research データ件数: 3,733 平均トークン数: 1,193 最大トークン数: 2,489 合計トークン数: 4,453,517 ファイル形式: JSONL ファイル分割数: 1 合計ファイルサイズ: 23.2 MB 加工内容: メタデータ列の解析と新列生成: metadata列(辞書型)を解析し、その中のreasoningをthought列に、difficultyをdifficulty列に展開しました。解析に失敗した行は除外されました。また、元のmetadata列は削除されました。 難易度によるフィルタリング:… See the full description on the dataset page: https://huggingface.co/datasets/LLMTeamAkiyama/cleand_moremilk_CoT_Reasoning_Scientific_Discovery_and_Research.tabularquestion-answering1K<n<10K0 likes11 downloads1y agoHugging Face22orion-research /gsmqnaoa-reasoning-pt_BRtext1K<n<10K1 likes9 downloads2y agoHugging Face23codin-research /sft-qa-synthetic-reasoninggatedtabular1K<n<10K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.