CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01caiovicentino1 /Qwen3.6-35B-A3B-mcr-stage-b Qwen3.6-35B-A3B — MCR Stage B Corpus (Distributed Reasoning Localization) First systematic mechanistic-intervention corpus on a hybrid MoE + GDN + Gated-Attention architecture. 📄 Paper: Loop-Intolerance Profiling: Localizing Distributed Reasoning in a Hybrid MoE Architecture via Nine Convergent Intervention Experiments — submitted to arXiv (2026-04-20, in moderation). Final arXiv ID will be added here once approved. This dataset contains per-token residual-stream activations at… See the full description on the dataset page: https://huggingface.co/datasets/caiovicentino1/Qwen3.6-35B-A3B-mcr-stage-b.textquestion-answeringn<1K1 likes1.2k downloads5mo agoHugging Face02katostrofik /qwen36-35b-a3b-fp8-two-blackhole-tt-cache Qwen3.6-35B-A3B-FP8 two-Blackhole TT cache This dataset contains the generated same-source compressed owner-bank cache used by a public Qwen/Qwen3.6-35B-A3B-FP8 two-Blackhole runtime project. Project repo: https://github.com/PMZFX/TT-qwen36-35b-a3b-fp8-two-blackhole The GitHub repo contains the runtime code, TT-Lang spike, reliability harnesses, release notes, and helper scripts. This dataset supplies the generated TT cache that is too large for the GitHub repo. Contents… See the full description on the dataset page: https://huggingface.co/datasets/katostrofik/qwen36-35b-a3b-fp8-two-blackhole-tt-cache.tabularn<1K0 likes286 downloads4mo agoHugging Face03phaedawg /qwen3.6-35b-a3b-distribution-fidelity-768x2048-v1 Qwen3.6-35B-A3B quantization analysis Mean KL divergence against on-disk size Scored under the distribution-fidelity laws, version 15. Read LAWS.md first: these numbers are comparable only within this artifact's token suite, geometry, and runtime identity, and not against any number produced elsewhere. Each candidate directory holds its one-pager (report.md), its raw report, its compliance receipt, and its Law 14 attribution where one was produced. reference/ carries the… See the full description on the dataset page: https://huggingface.co/datasets/phaedawg/qwen3.6-35b-a3b-distribution-fidelity-768x2048-v1.textn<1K0 likes280 downloads11d agoHugging Face04Edge0 /ark-asr-3b-open-asr-leaderboard-results ARK-ASR-3B Open ASR Leaderboard Results Raw JSONL manifests for AutoArk-AI/ARK-ASR-3B on the public English short-form hf-audio/open-asr-leaderboard splits. These manifests were generated on a local 8x RTX 4090 machine and scored with the shared Open ASR Leaderboard scorer: PYTHONPATH=. python - <<'PY' from normalizer.eval_utils import score_results score_results( 'ark_asr/results.AutoArk-AI-ARK-ASR-3B_20260622_official', 'AutoArk-AI/ARK-ASR-3B', ) PY Important:… See the full description on the dataset page: https://huggingface.co/datasets/Edge0/ark-asr-3b-open-asr-leaderboard-results.tabularautomatic-speech-recognition10K<n<100K12 likes185 downloads3mo agoHugging Face05zhuyksir /Ultrachat-Sharegpt-Qwen3-Next-80B-A3B-Instructtext100K<n<1M1 likes166 downloads1y agoHugging Face06twinkle-ai /llama-3.2-3B-f1-instruct-eval-logs-and-scorestabular100K<n<1M0 likes161 downloads7mo agoHugging Face07twinkle-ai /Llama-3.2-3B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes160 downloads7mo agoHugging Face08Efe2898 /Prosperity-Family-Alya-CPT-3B-Kumru Prosperity Family Alya CPT 3B — Kumru Tokenized Status: finished - 2B remain - 1B Developer: Prosperity AIModel/tokenizer: Efe2898/Prosperity-Family-Alya-BaseSource dataset: moganai/turkishfineweb2-cleaned Tokenizer Revision: df12a3c9e14c80d1464a8d7a2d624b2b021ab283Vocabulary: 50,176EOS token ID: 3 Filtering language_score >= 0.98 fasttext_clean_score >= 0.75 document chars: 200 .. 500000 deterministic stream shuffle seed: 20260906 shuffle buffer:… See the full description on the dataset page: https://huggingface.co/datasets/Efe2898/Prosperity-Family-Alya-CPT-3B-Kumru.tabularn<1K0 likes129 downloads17d agoHugging Face09ProCreations /grug-3b-train grug-3b-train training data for ProCreations/grug-3b. grug think in grug. grug answer in normal english. never other way round. what make this one different old grug model think short always. easy question, short think - good. hard question, short think - BAD. answer come out worse because grug not do the work. this set fix that. every fresh example carry difficulty tier, and tier decide how many word the think get. validator throw away think too short for tier… See the full description on the dataset page: https://huggingface.co/datasets/ProCreations/grug-3b-train.texttext-generation1K<n<10K5 likes103 downloads2mo agoHugging Face10bestdive /details_bestdive__SmolLM3-3B-SFT-Free-Course Smol course SFT evaluation - Kay Zheng Actual full GSM8K test evaluation of bestdive/SmolLM3-3B-SFT-Free-Course, adapter revision 0484e028b494d605a267050a949c9266edadd16b, merged with pinned SmolLM3-3B-Base before evaluation. Full 1319 test examples, zero-shot, original extractive_match: 0.4086429112964367 (stderr 0.013540639733342422). Free Google Colab T4, no paid HF Jobs; cost 0. lighteval 0.11.0, vLLM 0.10.1.1, Transformers 4.57.1, Python 3.12. Dataset-address correction… See the full description on the dataset page: https://huggingface.co/datasets/bestdive/details_bestdive__SmolLM3-3B-SFT-Free-Course.textn<1K0 likes99 downloads14d agoHugging Face11Phase-Technologies /forge-3b-dpo-data FORGE-3B DPO Preference Data Tokenized (prompt, chosen, rejected) preference triples for DPO post-training of FORGE-3B, built per the FORGE paper Section 6.2 / Appendix A.2. This is data preparation output only — no model was trained to produce this. Stats Total pairs: 0 (paper target: ~200,000) Domains: 0/4 Context length: 4096 tokens (paper Appendix A.2, DPO block) Format: unpacked — one (prompt, chosen, rejected) triple per training example Chat template:… See the full description on the dataset page: https://huggingface.co/datasets/Phase-Technologies/forge-3b-dpo-data.texttext-generation100K<n<1M0 likes94 downloads3mo agoHugging Face12visual-memory /ConvAI2-ERNIE-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-ernie-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes81 downloads14d agoHugging Face13visual-memory /ConvAI2-Qwen-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-qwen-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-Qwen-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-Qwen-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes80 downloads15d agoHugging Face143B-Group /ConvRetext10K<n<100K1 likes79 downloads2y agoHugging Face15JacobMolBio /vibethinker-3b-jlens-traces VibeThinker-3B J Lens Traces The J Lens is a Jacobian lens fitted to WeiboAI/VibeThinker-3B. Pick one of the 18 fitted source layers and a token position, and the lens decodes that residual-stream activation into a ranked list of vocabulary tokens. Following one position across layers shows how the decoded ranking changes on the way to the model's final output. This repository stores saved results from that lens. The companion model repository contains the lens weights. The… See the full description on the dataset page: https://huggingface.co/datasets/JacobMolBio/vibethinker-3b-jlens-traces.tabularn<1K0 likes77 downloads2mo agoHugging Face16visual-memory /ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-qwen-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes76 downloads14d agoHugging Face17visual-memory /ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-flux-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes73 downloads14d agoHugging Face18ericflo /Llama-3.2-3B-COTtext10K<n<100K0 likes70 downloads2y agoHugging Face19visual-memory /ConvAI2-FLUX-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-flux-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-FLUX-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-FLUX-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes69 downloads14d agoHugging Face20visual-memory /ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-ernie-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes66 downloads14d agoHugging Face21LeeXugar /CodePin-SFT-Qwen3.5-35B-A3B CodePin SFT — Qwen3.5-35B-A3B Teacher Trajectories CodePin SFT contains 6,000 validated code-localization trajectories generated with qwen3.5-35b-a3b. It is designed for pure-text supervised fine-tuning of Qwen/Qwen3.5-0.8B and other tool-calling language models. The tasks come from LeeXugar/SWE-smith-code-search. The source rollout dataset was used only for cleaning, difficulty estimation, and sample selection; rollout messages and rewards were not copied into these SFT… See the full description on the dataset page: https://huggingface.co/datasets/LeeXugar/CodePin-SFT-Qwen3.5-35B-A3B.texttext-generation1K<n<10K0 likes65 downloads1mo agoHugging Face22Hsueh008 /Qwen2.5-3B-jsonltabular100K<n<1M0 likes63 downloads1mo agoHugging Face23lhchau /llama3.2-3B-sfttext10K<n<100K0 likes62 downloads3mo agoHugging Face24open-llm-leaderboard /bigcode__starcoder2-3b-detailsgated Dataset Card for Evaluation run of bigcode/starcoder2-3b Dataset automatically created during the evaluation run of model bigcode/starcoder2-3b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigcode__starcoder2-3b-details.tabular10K<n<100K0 likes61 downloads2y agoHugging Face25open-llm-leaderboard /NousResearch__Hermes-3-Llama-3.2-3B-detailsgated Dataset Card for Evaluation run of NousResearch/Hermes-3-Llama-3.2-3B Dataset automatically created during the evaluation run of model NousResearch/Hermes-3-Llama-3.2-3B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Hermes-3-Llama-3.2-3B-details.tabular10K<n<100K0 likes61 downloads2y agoHugging Face26visual-memory /PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-flux-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes58 downloads15d agoHugging Face27visual-memory /PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-ernie-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes58 downloads15d agoHugging Face28visual-memory /PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-qwen-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes56 downloads15d agoHugging Face29baohao /Math_Qwen3-30B-A3B-Instruct-2507_SFT-RL_Prefixtext10K<n<100K1 likes55 downloads2mo agoHugging Face30SeanWang0027 /polaris_rose_rollouts_olmo3-7b_from_qwen3-30b-a3b_cutoff4096_240steps Cross-tokenizer ROSE rollouts — Olmo-3-7B-Think-SFT ← Qwen3-30B-A3B-Thinking-2507 Every assembled row of a complete 240-step online-ROSE run: 61,440 rows, the teacher's actual continuation for each, and the token accounting behind it. The student writes a 4096-token prefix in its own vocabulary (100278). That prefix is decoded to text, the teacher is shown it under its own chat template, and the teacher's reply comes back as text and is tokenised into the student's vocabulary.… See the full description on the dataset page: https://huggingface.co/datasets/SeanWang0027/polaris_rose_rollouts_olmo3-7b_from_qwen3-30b-a3b_cutoff4096_240steps.tabulartext-generation10K<n<100K0 likes54 downloads25d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.