CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SaylorTwift /RULER-8192-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes1.6k downloads1y agoHugging Face02ENSEONG /full-math-private-n256-Qwen2.5-3B-Instruct-bontabular100K<n<1M0 likes1.4k downloads6mo agoHugging Face03ENSEONG /full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes1.1k downloads5mo agoHugging Face04ENSEONG /preprocessed-full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes944 downloads5mo agoHugging Face05ENSEONG /full-math-private-Qwen2.5-3B-Instruct-bontabular100K<n<1M0 likes777 downloads6mo agoHugging Face06ENSEONG /stratified-solvable-1k-math-private-Qwen2.5-3B-Instruct-bontabular10K<n<100K0 likes692 downloads6mo agoHugging Face07marin-community /openthoughts4-code-9168-prompts-qwen3-30b-a3b-thinking-2507-n16-flattened-logprobs-k16 OpenThoughts-4 Code SDG: Qwen3-30B-A3B-Thinking-2507 (n=16, top-16 logprobs) Synthetic generations from Qwen/Qwen3-30B-A3B-Thinking-2507 on the Marin OpenThoughts-4 code SDG prompt set. Each prompt is sampled n=16 times, and for every generated token the dataset stores the chosen-token log probability plus the top-16 log probabilities over the vocabulary, enabling distillation, KL-style fine-tuning, reranking, and uncertainty analysis. Generation setup Field… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/openthoughts4-code-9168-prompts-qwen3-30b-a3b-thinking-2507-n16-flattened-logprobs-k16.tabulartext-generation100K<n<1M0 likes595 downloads5mo agoHugging Face08juiceb0xc0de /AI21-Jamba2-3B juiceb0xc0de/AI21-Jamba2-3B A brain atlas for ai21labs/AI21-Jamba2-3B, a 28-layer hybrid Mamba/transformer from AI21 Labs. This is not a chat dataset or a benchmark. It is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and direction is doing. Jamba is an interesting subject because it is mostly not attention. Of the 28 layers, only 2 carry attention, and both of those run a single KV head.… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/AI21-Jamba2-3B.imagefeature-extraction1M<n<10M0 likes591 downloads7d agoHugging Face09rl-rag /browsecomp-qwen35-35b-a3b-think browsecomp-qwen35-35b-a3b-think Deep research agent evaluation on data/browsecomp.jsonl (normal split). Results Metric Value pass@4 43.0% avg@4 24.8% Trajectory accuracy 24.8% (1258/5064) Questions 1266 Trajectories 5064 (4 per question) Avg tool calls 41.1 Full conversations ❌ Model & Setup Model Qwen3.5-35B-A3B Judge gpt-4o Max tool calls 50 Temperature 0.7 Blocked domains huggingface.co Tool… See the full description on the dataset page: https://huggingface.co/datasets/rl-rag/browsecomp-qwen35-35b-a3b-think.tabular1K<n<10K0 likes573 downloads7mo agoHugging Face10kothasuhas /llama-3b-gold-15M-student-generations_SNIS_2048_tune422v1tabular10M<n<100M0 likes521 downloads1y agoHugging Face11SaylorTwift /RULER-32768-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes500 downloads1y agoHugging Face12Lansechen /details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-default Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-default Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-default. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-default.tabular1K<n<10K0 likes449 downloads1y agoHugging Face13ftajwar /uniagent-qwen3-30b-a3b-r2e-rolloutstabular100K<n<1M0 likes439 downloads11d agoHugging Face14hanlincs /in1k_clip_qwen25vl_3b_224res_64tokens_new_pttabular1M<n<10M0 likes409 downloads1y agoHugging Face15hanlincs /in1k_clip_qwen25vl_3b_448res_256tokens_new_merged_pttabular1M<n<10M0 likes394 downloads1y agoHugging Face16kothasuhas /llama-3b-gold-15M-student-generations_PRESAMPLING_2048_tune422v1tabular10M<n<100M0 likes382 downloads1y agoHugging Face17marin-community /openthoughts4-science-26041-prompts-qwen3-30b-a3B-thinking-2507-n8-flattened-logprobs-k16 OpenThoughts-4 Science SDG: Qwen3-30B-A3B-Thinking-2507 (n=8, top-16 logprobs) Synthetic generations from Qwen/Qwen3-30B-A3B-Thinking-2507 on the Marin OpenThoughts-4 science SDG prompt set. Each prompt is sampled n=8 times, and for every generated token the dataset stores the chosen-token log probability plus the top-16 log probabilities over the vocabulary, enabling distillation, KL-style fine-tuning, reranking, and uncertainty analysis. Generation setup Field… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/openthoughts4-science-26041-prompts-qwen3-30b-a3B-thinking-2507-n8-flattened-logprobs-k16.tabulartext-generation100K<n<1M0 likes351 downloads5mo agoHugging Face18katostrofik /qwen36-35b-a3b-fp8-two-blackhole-tt-cache Qwen3.6-35B-A3B-FP8 two-Blackhole TT cache This dataset contains the generated same-source compressed owner-bank cache used by a public Qwen/Qwen3.6-35B-A3B-FP8 two-Blackhole runtime project. Project repo: https://github.com/PMZFX/TT-qwen36-35b-a3b-fp8-two-blackhole The GitHub repo contains the runtime code, TT-Lang spike, reliability harnesses, release notes, and helper scripts. This dataset supplies the generated TT cache that is too large for the GitHub repo. Contents… See the full description on the dataset page: https://huggingface.co/datasets/katostrofik/qwen36-35b-a3b-fp8-two-blackhole-tt-cache.tabularn<1K0 likes333 downloads4mo agoHugging Face19juiceb0xc0de /llama-3.2-3b-atlas llama-3.2-3b-atlas image1M<n<10M0 likes313 downloads25d agoHugging Face20OALL /details_Qwen__Qwen3-30B-A3B-Thinking-2507_v2 Dataset Card for Evaluation run of Qwen/Qwen3-30B-A3B-Thinking-2507 Dataset automatically created during the evaluation run of model Qwen/Qwen3-30B-A3B-Thinking-2507. The dataset is composed of 116 configuration, each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Qwen__Qwen3-30B-A3B-Thinking-2507_v2.tabular100K<n<1M0 likes305 downloads8mo agoHugging Face21ENSEONG /preprocessed-full-math-private-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes293 downloads6mo agoHugging Face22mlfoundations-dev /DeepHermes-3-Llama-3-3B-Preview_eval_2e29 mlfoundations-dev/DeepHermes-3-Llama-3-3B-Preview_eval_2e29 Precomputed model outputs for evaluation. Evaluation Results Summary Metric AIME24 AMC23 MATH500 MMLUPro JEEBench GPQADiamond LiveCodeBench CodeElo CodeForces AIME25 HLE LiveCodeBenchv5 Accuracy 0.0 2.5 5.2 18.2 3.5 2.7 2.3 1.4 3.8 0.0 8.0 1.5 AIME24 Average Accuracy: 0.00% ± 0.00% Number of Runs: 10 Run Accuracy Questions Solved Total Questions 1 0.00% 0 30… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/DeepHermes-3-Llama-3-3B-Preview_eval_2e29.tabular10K<n<100K0 likes289 downloads1y agoHugging Face23ENSEONG /full-gsm8k-private-n256-Qwen2.5-3B-Instruct-bontabular10K<n<100K0 likes254 downloads5mo agoHugging Face24ENSEONG /full-math-private-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes253 downloads6mo agoHugging Face25juiceb0xc0de /llama3.2-3b-instruct-atlas juiceb0xc0de/llama3.2-3b-instruct-atlas A brain atlas for meta-llama/Llama-3.2-3B-Instruct, the 3B instruction-tuned member of the Llama 3.2 family. This is not a chat dataset or a benchmark - it is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing. If you want to know where an instruction-tuned model keeps its register machinery, which directions survive a causal test… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama3.2-3b-instruct-atlas.imagefeature-extraction1M<n<10M0 likes247 downloads25d agoHugging Face26PatrickHaller /fineweb-edu-3Btabular1M<n<10M0 likes240 downloads2y agoHugging Face27ENSEONG /preprocessed-full-aime_2023-n256-Qwen2.5-3B-Instruct-bontabular1K<n<10K0 likes226 downloads4mo agoHugging Face28marin-community /open-thoughts-4-30k-code-qwen3-30b-a3B-thinking-2507-annotated-32768-tokens-n8 Open Thoughts 4 - Code (Qwen3-30B-A3B-Thinking-2507, 32K tokens, n=8) This dataset contains code reasoning problems with 8 independent responses generated by Qwen3-30B-A3B-Thinking-2507. Overview Source: marin-community/open-thoughts-4-30k-code-qwen3-32b-annotated (prompts only) Model: Qwen/Qwen3-30B-A3B-Thinking-2507 Temperature: 0.8 Max tokens: 32,768 Columns Column Description instruction_seed The code problem prompt _source Source dataset… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-code-qwen3-30b-a3B-thinking-2507-annotated-32768-tokens-n8.tabular10K<n<100K0 likes217 downloads7mo agoHugging Face29ENSEONG /preprocessed-full-gsm8k-private-n256-Qwen2.5-3B-Instruct-bontabular10K<n<100K0 likes212 downloads5mo agoHugging Face30Lansechen /details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-cosine-v2 Dataset Card for Evaluation run of Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-cosine-v2 Dataset automatically created during the evaluation run of model Lansechen/Qwen2.5-3B-Open-R1-GRPO-math-selected-cosine-v2. The dataset is composed of 3 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 10 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/Lansechen/details_Lansechen__Qwen2.5-3B-Open-R1-GRPO-math-selected-cosine-v2.tabular1K<n<10K0 likes205 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.