CoolFace
17 results

research-reasoning

mlfoundations-dev /Nemotron-Research-Reasoning-Qwen-1.5B_eval_569atabular1K<n<10K0 likes1.1k downloads1y agoHugging FaceAmanPriyanshu /tool-reasoning-sft-RESEARCH-openresearcher-dataset-sft-deep-research-agent-data-cleaned OpenResearcher Dataset - Cleaned & Restructured 👥 Follow the Author Aman Priyanshu Overview This dataset is a cleaned and restructured version of the OpenResearcher Dataset released by the TIGER-AI-Lab. The original dataset contains 96K+ long-horizon deep research trajectories generated by GPT-OSS-120B with native browser tools. This version converts the GPT-OSS channel-based message format into a standardized multi-turn tool-use conversation… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-openresearcher-dataset-sft-deep-research-agent-data-cleaned.text-generation10K<n<100K2 likes294 downloads6mo agoHugging FaceKylan12 /mycotoxin-chemical-research-sythetic-reasoning mycotoxin-chemical-research-sythetic-reasoning Synthetic Q&A dataset on Mycotoxin Chemical Research, generated with SDGS (Synthetic Dataset Generation Suite). Dataset Details Metric Value Topic Mycotoxin Chemical Research Total Q&A Pairs 4416 Valid Pairs 4416 Provider/Model ollama/gpt-oss:120b Generation Cost Metric Value Prompt Tokens 4,579,253 Completion Tokens 5,326,284 Total Tokens 9,905,537 GPU Energy 3.5712 kWh… See the full description on the dataset page: https://huggingface.co/datasets/Kylan12/mycotoxin-chemical-research-sythetic-reasoning.textquestion-answering1K<n<10K0 likes198 downloads7mo agoHugging Faceulamai /verified-research-reasoning-trajectories Verified Research Reasoning Trajectories for RLVR This repository is the public sample and schema repository for Ulam's research-level mathematical reasoning trajectories for reinforcement learning with verifiable rewards (RLVR), process supervision, judge training, proof criticism, and private evaluations. Ulam Verified Research Reasoning Trajectories are proof-process data for RLVR. Each record contains a normalized research problem, a golden or partial-golden proof graph… See the full description on the dataset page: https://huggingface.co/datasets/ulamai/verified-research-reasoning-trajectories.documenttext-generationn<1K3 likes174 downloads2mo agoHugging Faceoof-baroomf /reasoning-switch-research-artifacts0 likes139 downloads2mo agoHugging Facereasoning-degeneration-dev /RESEARCH_DASHBOARD0 likes131 downloads6mo agoHugging Face