CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alexbouayad /stack-v2-starcoder2-3btext100K<n<1M0 likes2.4k downloads10d agoHugging Face02SaylorTwift /RULER-8192-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes1.6k downloads1y agoHugging Face03ENSEONG /full-math-private-n256-Qwen2.5-3B-Instruct-bontabular100K<n<1M0 likes1.4k downloads6mo agoHugging Face04mlfoundations-cua-dev /easyr1-grounding-dataset-30k-not_grounded-SE-GUI-3B-2MPimage10K<n<100K1 likes1.3k downloads1y agoHugging Face05caiovicentino1 /Qwen3.6-35B-A3B-mcr-stage-b Qwen3.6-35B-A3B — MCR Stage B Corpus (Distributed Reasoning Localization) First systematic mechanistic-intervention corpus on a hybrid MoE + GDN + Gated-Attention architecture. 📄 Paper: Loop-Intolerance Profiling: Localizing Distributed Reasoning in a Hybrid MoE Architecture via Nine Convergent Intervention Experiments — submitted to arXiv (2026-04-20, in moderation). Final arXiv ID will be added here once approved. This dataset contains per-token residual-stream activations at… See the full description on the dataset page: https://huggingface.co/datasets/caiovicentino1/Qwen3.6-35B-A3B-mcr-stage-b.textquestion-answeringn<1K1 likes1.2k downloads5mo agoHugging Face06ENSEONG /full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes1.1k downloads5mo agoHugging Face07zake7749 /Qwen3.6-35B-A3B-Tool-Calling Qwen3.6-35B-A3B Tool-Calling Dataset This repository presents a function and tool-calling preference and supervised fine-tuning dataset constructed from Nemotron-RL agentic prompt corpora. For each source prompt, the model was sampled four times with thinking mode enabled. Each resulting candidate trajectory was then evaluated against the dataset’s ground-truth action using exact matching on both the function name and the parsed function arguments. Overview… See the full description on the dataset page: https://huggingface.co/datasets/zake7749/Qwen3.6-35B-A3B-Tool-Calling.text10K<n<100K15 likes1.1k downloads5mo agoHugging Face08ENSEONG /preprocessed-full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes944 downloads5mo agoHugging Face09ENSEONG /full-math-private-Qwen2.5-3B-Instruct-bontabular100K<n<1M0 likes777 downloads6mo agoHugging Face10ENSEONG /stratified-solvable-1k-math-private-Qwen2.5-3B-Instruct-bontabular10K<n<100K0 likes692 downloads6mo agoHugging Face11laion /terminal_bench_2_tasktrove_dq_stack_pytest_step25_30b_a3b_20260730_053956 TaskTrove stack-pytest — training rollout traces (Qwen3-Coder-30B-A3B, step 25) Terminus-2/Harbor rollouts recorded while training laion/tasktrove-dq-stack-pytest-step25-30b-a3b with SkyRL on the TaskTrove stack-pytest source. One row per trial, holding that trial's last episode as an OpenAI-style conversations list, the task instruction, the reward the verifier assigned (result), and the verifier's own stdout (verifier_output). Source run… See the full description on the dataset page: https://huggingface.co/datasets/laion/terminal_bench_2_tasktrove_dq_stack_pytest_step25_30b_a3b_20260730_053956.text10K<n<100K0 likes626 downloads2mo agoHugging Face12marin-community /openthoughts4-code-9168-prompts-qwen3-30b-a3b-thinking-2507-n16-flattened-logprobs-k16 OpenThoughts-4 Code SDG: Qwen3-30B-A3B-Thinking-2507 (n=16, top-16 logprobs) Synthetic generations from Qwen/Qwen3-30B-A3B-Thinking-2507 on the Marin OpenThoughts-4 code SDG prompt set. Each prompt is sampled n=16 times, and for every generated token the dataset stores the chosen-token log probability plus the top-16 log probabilities over the vocabulary, enabling distillation, KL-style fine-tuning, reranking, and uncertainty analysis. Generation setup Field… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/openthoughts4-code-9168-prompts-qwen3-30b-a3b-thinking-2507-n16-flattened-logprobs-k16.tabulartext-generation100K<n<1M0 likes595 downloads5mo agoHugging Face13juiceb0xc0de /AI21-Jamba2-3B juiceb0xc0de/AI21-Jamba2-3B A brain atlas for ai21labs/AI21-Jamba2-3B, a 28-layer hybrid Mamba/transformer from AI21 Labs. This is not a chat dataset or a benchmark. It is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and direction is doing. Jamba is an interesting subject because it is mostly not attention. Of the 28 layers, only 2 carry attention, and both of those run a single KV head.… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/AI21-Jamba2-3B.imagefeature-extraction1M<n<10M0 likes591 downloads7d agoHugging Face14rl-rag /browsecomp-qwen35-35b-a3b-think browsecomp-qwen35-35b-a3b-think Deep research agent evaluation on data/browsecomp.jsonl (normal split). Results Metric Value pass@4 43.0% avg@4 24.8% Trajectory accuracy 24.8% (1258/5064) Questions 1266 Trajectories 5064 (4 per question) Avg tool calls 41.1 Full conversations ❌ Model & Setup Model Qwen3.5-35B-A3B Judge gpt-4o Max tool calls 50 Temperature 0.7 Blocked domains huggingface.co Tool… See the full description on the dataset page: https://huggingface.co/datasets/rl-rag/browsecomp-qwen35-35b-a3b-think.tabular1K<n<10K0 likes573 downloads7mo agoHugging Face15Hzfinfdu /SlimPajama-3Btext1M<n<10M2 likes550 downloads2y agoHugging Face16laion /terminal_bench_2_tasktrove_dq_unix_step10_30b_a3b_20260730_014756 terminal_bench_2_tasktrove_dq_unix_step10_30b_a3b OpenCode agent trajectories from the TaskTrove DQ unix arm of a Qwen3-Coder-30B-A3B agentic RL sweep, exported from the complete Harbor rollout artifact set. Coverage Built from the full trace_jobs prefix of run rl-tasktrove-dq-sweep-30b-qwen3-coder-30-20260727-082204-e42f1d (12034 trial directories, 11937 of them scored). quantity value scored trials (result.json) 11937 rows published 11937 coverage… See the full description on the dataset page: https://huggingface.co/datasets/laion/terminal_bench_2_tasktrove_dq_unix_step10_30b_a3b_20260730_014756.text10K<n<100K0 likes531 downloads2mo agoHugging Face17kothasuhas /llama-3b-gold-15M-student-generations_SNIS_2048_tune422v1tabular10M<n<100M0 likes521 downloads1y agoHugging Face18kothasuhas /llama-3b-gold-15M-student-generationstext10M<n<100M0 likes517 downloads1y agoHugging Face19SaylorTwift /RULER-32768-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes500 downloads1y agoHugging Face20cheesewafer /qwen3.6-35B-A3B_resultsimagen<1K0 likes491 downloads7d agoHugging Face21nishadsinghi /MATH_train_llama3.2-3b-instructtext1K<n<10K0 likes463 downloads2y agoHugging Face22llm-jp /llm-jp-4-32b-a3b-thinking-dpo-data llm-jp-4-32b-a3b-thinking-dpo-data Overview This dataset is a Direct Preference Optimization (DPO) dataset used to train llm-jp-4-32b-a3b-thinking. It is constructed by pairing multiple candidate responses for a given prompt and selecting preferred (chosen) and non-preferred (rejected) responses. The splits reasoning_low, reasoning_medium, and reasoning_high correspond to different reasoning effort settings used during response generation. The fields chosen_analysis… See the full description on the dataset page: https://huggingface.co/datasets/llm-jp/llm-jp-4-32b-a3b-thinking-dpo-data.text100K<n<1M1 likes457 downloads5mo agoHugging Face23Lansechen /details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-default Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-default Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-default. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-default.tabular1K<n<10K0 likes449 downloads1y agoHugging Face24ftajwar /uniagent-qwen3-30b-a3b-r2e-rolloutstabular100K<n<1M0 likes439 downloads11d agoHugging Face25sehyun734 /longhealth-llama-3.2-3btext10K<n<100K1 likes430 downloads4d agoHugging Face26michaelm16 /GuideRNA-3B Dataset Card for GuideRNA-3B Dataset Summary GuideRNA-3B is a large transcriptome sequence corpus consisting of over 3.7 billion paired sequences extracted from the specific transcriptome of 23 cell lines and over 200 segmented genomes of RNA virus. Supported Tasks Based on this nucleotide sequence corpus, we are able to establish a foundation model to characterize the manifold of CRISPR guide RNA targeting regions in order to undertake further downstreaming… See the full description on the dataset page: https://huggingface.co/datasets/michaelm16/GuideRNA-3B.text1B<n<10B1 likes414 downloads2y agoHugging Face27hanlincs /in1k_clip_qwen25vl_3b_224res_64tokens_new_pttabular1M<n<10M0 likes409 downloads1y agoHugging Face28hanlincs /in1k_clip_qwen25vl_3b_448res_256tokens_new_merged_pttabular1M<n<10M0 likes394 downloads1y agoHugging Face29nanoswe /swesmith-qwen3.6-35b-a3b SWE-smith trajectories from Qwen3.6-35B-A3B Multi-turn coding-agent trajectories (issue → tool-using rollout → patch) produced by Qwen3.6-35B-A3B on SWE-smith tasks, stored untokenized. This is the exact SFT corpus used for the harbor arm of the nanoswe teacher-distillation experiments. 101,901 trajectories over 45,242 unique SWE-smith task instances (3 sampled rollouts per task, ~2.25 surviving filtering), 53 parquet shards, ~1.4 GB. ≈1.96B training tokens = exactly one epoch… See the full description on the dataset page: https://huggingface.co/datasets/nanoswe/swesmith-qwen3.6-35b-a3b.texttext-generation100K<n<1M0 likes389 downloads1mo agoHugging Face30kothasuhas /llama-3b-gold-15M-student-generations_PRESAMPLING_2048_tune422v1tabular10M<n<100M0 likes382 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.