CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01rl-research /deep_research_bench_evaltextn<1K0 likes1.2k downloads10mo agoHugging Face02kylemontgomery /deepresearch-corpustext100K<n<1M0 likes1.1k downloads5mo agoHugging Face03ScienceOne-AI /S1-DeepResearch-15k S1-DeepResearch-15k Dataset Overview The S1-DeepResearch dataset is a curated collection of approximately 15k samples designed to improve deep research capabilities of large language models. The dataset includes two types of tasks: Verifiable tasks (labeled as "Closed-ended Multi-hop Resolution") Open-ended tasks (labeled as "Open-ended Exploration") Dataset Composition The dataset is organized into five core capability dimensions: Long-chain complex… See the full description on the dataset page: https://huggingface.co/datasets/ScienceOne-AI/S1-DeepResearch-15k.text10K<n<100K12 likes674 downloads5mo agoHugging Face04ninja-x /deepresearchtextn<1K0 likes324 downloads2y agoHugging Face05HuanjinYao /MM-DeepResearch-corpustext1K<n<10K0 likes312 downloads5mo agoHugging Face06ICA-DeepResearch /hermes-bc-traj Browse Comp Eval Standalone parallel Hermes runner. It does not import AIDABench; it only follows the same operational shape: JSONL input, concurrent runs, retry, resume, and per-run JSON outputs. Input Put JSONL files under data/{dataset}/. Each row must contain: {"question": "...", "answer": "...", "type": "..."} answer is only recorded for later evaluation. It is not sent to Hermes. Run cd /root/Browse_comp_eval export HERMES_API_KEY="..."… See the full description on the dataset page: https://huggingface.co/datasets/ICA-DeepResearch/hermes-bc-traj.documentn<1K0 likes290 downloads2mo agoHugging Face07IPF /DeepResearch-traj DeepResearch-traj Multi-seed deep research agent trajectories with per-question correctness labels and pass@k statistics, derived from OpenResearcher/OpenResearcher-Dataset. Dataset Summary This dataset contains 97,630 full agent trajectories across 6,102 unique research questions, each sampled under 16 different random seeds (42–57). Every trajectory is annotated with: seed — which random seed produced this trajectory correct — whether the model's final answer was… See the full description on the dataset page: https://huggingface.co/datasets/IPF/DeepResearch-traj.tabularquestion-answering10K<n<100K0 likes287 downloads7mo agoHugging Face08artillerywu /DeepResearch-9K Data Splits The dataset consists of the following two subsets: Dataset Name Contents How It Was Generated Number of Samples DeepResearch-9K All samples (teacher model's outputs) Teacher model inference on 9K questions 9,000 DeepResearch-Hard Teacher model's incorrect samples only Filtered from DeepResearch-9K (samples where the teacher model's final answer was wrong) 3,974 Based on the above, we define the following train/test split for our DeepResearch-R1 model:… See the full description on the dataset page: https://huggingface.co/datasets/artillerywu/DeepResearch-9K.text10K<n<100K18 likes267 downloads5mo agoHugging Face09allenai /deepresearch-benchtextn<1K0 likes242 downloads2mo agoHugging Face10cx-cmu /deepresearchgym-agentic-search-logs DeepResearchGym Agentic Search Logs This repository hosts the dataset accompanying the paper “Agentic Search in the Wild” (arXiv: https://arxiv.org/abs/2601.17617). The dataset contains 14M+ search queries collected via DeepResearchGym (DRGym), an open-source search API designed for DeepResearch-style agentic search. For more background on DRGym, see: https://arxiv.org/abs/2505.19253. All records have been anonymized and shuffled to prevent re-identification, and we additionally… See the full description on the dataset page: https://huggingface.co/datasets/cx-cmu/deepresearchgym-agentic-search-logs.tabulartext-retrieval10M<n<100M16 likes237 downloads8mo agoHugging Face11lee64 /deepresearch-bench-querytextn<1K0 likes208 downloads11mo agoHugging Face12InternScience /SGI-DeepResearchgated Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows &nbsp; &nbsp; &nbsp; Welcome to the official repository for the SGI-Bench! 👏 Scientist-aligned benchmark for evaluating Scientific General Intelligence (SGI) across the full inquiry cycle: Deliberation, Conception, Action, and Perception. The benchmark spans 10 disciplines and more than 1,000 expert‑curated samples inspired by Science’s 125 Big Questions, with an agentic evaluation framework… See the full description on the dataset page: https://huggingface.co/datasets/InternScience/SGI-DeepResearch.textquestion-answeringn<1K11 likes146 downloads4mo agoHugging Face13JRQi /DeepResearch-Bench-Multilingual DeepResearch Bench Multilingual Prompts This dataset provides prompt-level multilingual translations for the 100 research tasks used in muset-ai/DeepResearch-Bench-Dataset. The translations cover eight languages: en zh es it ar bn ja el What is included This repository focuses on the benchmark prompts only. On the Hugging Face Hub, the Dataset Viewer is configured with one default subset named all plus nine explicit subset configurations: source_prompt, en, zh, es… See the full description on the dataset page: https://huggingface.co/datasets/JRQi/DeepResearch-Bench-Multilingual.texttext-generation1K<n<10K1 likes120 downloads6mo agoHugging Face14kizro /deep_research_taskset_fulltextn<1K0 likes115 downloads1y agoHugging Face15ninja-x /deepresearch-v1textn<1K1 likes112 downloads2y agoHugging Face16Osilly /Vision-DeepResearch-Evaltext1K<n<10K0 likes105 downloads4mo agoHugging Face17Deep-Research-Team /Pre-Training-Persian-Corpus-Raw-Texts-DatasetDocument Version: 2.0.0 | Last Updated: 02/13/2026 text10M<n<100M0 likes98 downloads7mo agoHugging Face18FractalAIResearch /DeepResearch-SFT Fathom-DeepResearch: Unlocking Long Horizon Information Retrieval And Synthesis For SLMs ✨ News [29/09/25]: Our paper on Fathom-Search-4B has been accepted to SEA @ NeurIPS 2025 🎉 OpenReview link Introduction We introduce Fathom-DeepResearch, an agentic DeepResearch system that sets state-of-the-art performance in the open-weights category on search-intensive benchmarks (SimpleQA, FRAMES, WebWalkerQA, Seal0) and outperforms… See the full description on the dataset page: https://huggingface.co/datasets/FractalAIResearch/DeepResearch-SFT.text1K<n<10K11 likes96 downloads1y agoHugging Face19deepresearchediting /deer_sample_queriestextn<1K0 likes94 downloads6mo agoHugging Face20NovaSky-AI /DeepResearch-Datatext1K<n<10K1 likes90 downloads10mo agoHugging Face21kizro /deep_research_tasksettextn<1K0 likes88 downloads1y agoHugging Face22kylemontgomery /deep-research-sft-0406text10K<n<100K1 likes80 downloads6mo agoHugging Face23CostaliyA /Vision-DeepResearch-Text-Datatext1K<n<10K0 likes74 downloads1mo agoHugging Face24ICA-DeepResearch /ICA-SFT-14ktext10K<n<100K1 likes73 downloads4mo agoHugging Face25Osilly /Vision-DeepResearch-Toy-SFT-Datatext1K<n<10K0 likes66 downloads8mo agoHugging Face26Infinity-AILab /DeepResearchEvalThis dataset contains 100 high-quality deep research tasks from DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation. GitHub repository: https://github.com/Infinity-AILab/DeepResearchEval textn<1K0 likes65 downloads9mo agoHugging Face27kylemontgomery /deepresearch-taskstabular100K<n<1M1 likes63 downloads7mo agoHugging Face28Osilly /Vision-DeepResearch-Toy-RL-Datatext1K<n<10K0 likes57 downloads8mo agoHugging Face29erobinson /repro-mm-deepresearch-a-simple-and-effective-multimodal-agentic-search-baseline-traces Agent traces Agent sessions published from a Trackio Logbook. textn<1K0 likes55 downloads2mo agoHugging Face30kizro /deep_research_taskset_full_filteredtextn<1K0 likes53 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.