CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01caiovicentino1 /Qwen3.6-35B-A3B-mcr-stage-b Qwen3.6-35B-A3B — MCR Stage B Corpus (Distributed Reasoning Localization) First systematic mechanistic-intervention corpus on a hybrid MoE + GDN + Gated-Attention architecture. 📄 Paper: Loop-Intolerance Profiling: Localizing Distributed Reasoning in a Hybrid MoE Architecture via Nine Convergent Intervention Experiments — submitted to arXiv (2026-04-20, in moderation). Final arXiv ID will be added here once approved. This dataset contains per-token residual-stream activations at… See the full description on the dataset page: https://huggingface.co/datasets/caiovicentino1/Qwen3.6-35B-A3B-mcr-stage-b.textquestion-answeringn<1K1 likes1.2k downloads5mo agoHugging Face02katostrofik /qwen36-35b-a3b-fp8-two-blackhole-tt-cache Qwen3.6-35B-A3B-FP8 two-Blackhole TT cache This dataset contains the generated same-source compressed owner-bank cache used by a public Qwen/Qwen3.6-35B-A3B-FP8 two-Blackhole runtime project. Project repo: https://github.com/PMZFX/TT-qwen36-35b-a3b-fp8-two-blackhole The GitHub repo contains the runtime code, TT-Lang spike, reliability harnesses, release notes, and helper scripts. This dataset supplies the generated TT cache that is too large for the GitHub repo. Contents… See the full description on the dataset page: https://huggingface.co/datasets/katostrofik/qwen36-35b-a3b-fp8-two-blackhole-tt-cache.tabularn<1K0 likes333 downloads4mo agoHugging Face03phaedawg /qwen3.6-35b-a3b-distribution-fidelity-768x2048-v1 Qwen3.6-35B-A3B quantization analysis Mean KL divergence against on-disk size Scored under the distribution-fidelity laws, version 15. Read LAWS.md first: these numbers are comparable only within this artifact's token suite, geometry, and runtime identity, and not against any number produced elsewhere. Each candidate directory holds its one-pager (report.md), its raw report, its compliance receipt, and its Law 14 attribution where one was produced. reference/ carries the… See the full description on the dataset page: https://huggingface.co/datasets/phaedawg/qwen3.6-35b-a3b-distribution-fidelity-768x2048-v1.textn<1K0 likes280 downloads10d agoHugging Face04Edge0 /ark-asr-3b-open-asr-leaderboard-results ARK-ASR-3B Open ASR Leaderboard Results Raw JSONL manifests for AutoArk-AI/ARK-ASR-3B on the public English short-form hf-audio/open-asr-leaderboard splits. These manifests were generated on a local 8x RTX 4090 machine and scored with the shared Open ASR Leaderboard scorer: PYTHONPATH=. python - <<'PY' from normalizer.eval_utils import score_results score_results( 'ark_asr/results.AutoArk-AI-ARK-ASR-3B_20260622_official', 'AutoArk-AI/ARK-ASR-3B', ) PY Important:… See the full description on the dataset page: https://huggingface.co/datasets/Edge0/ark-asr-3b-open-asr-leaderboard-results.tabularautomatic-speech-recognition10K<n<100K12 likes195 downloads3mo agoHugging Face05zhuyksir /Ultrachat-Sharegpt-Qwen3-Next-80B-A3B-Instructtext100K<n<1M1 likes193 downloads1y agoHugging Face06twinkle-ai /llama-3.2-3B-f1-instruct-eval-logs-and-scorestabular100K<n<1M0 likes156 downloads7mo agoHugging Face07twinkle-ai /Llama-3.2-3B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes156 downloads7mo agoHugging Face08Phase-Technologies /forge-3b-dpo-data FORGE-3B DPO Preference Data Tokenized (prompt, chosen, rejected) preference triples for DPO post-training of FORGE-3B, built per the FORGE paper Section 6.2 / Appendix A.2. This is data preparation output only — no model was trained to produce this. Stats Total pairs: 0 (paper target: ~200,000) Domains: 0/4 Context length: 4096 tokens (paper Appendix A.2, DPO block) Format: unpacked — one (prompt, chosen, rejected) triple per training example Chat template:… See the full description on the dataset page: https://huggingface.co/datasets/Phase-Technologies/forge-3b-dpo-data.texttext-generation100K<n<1M0 likes138 downloads3mo agoHugging Face09Efe2898 /Prosperity-Family-Alya-CPT-3B-Kumru Prosperity Family Alya CPT 3B — Kumru Tokenized Status: finished - 2B remain - 1B Developer: Prosperity AIModel/tokenizer: Efe2898/Prosperity-Family-Alya-BaseSource dataset: moganai/turkishfineweb2-cleaned Tokenizer Revision: df12a3c9e14c80d1464a8d7a2d624b2b021ab283Vocabulary: 50,176EOS token ID: 3 Filtering language_score >= 0.98 fasttext_clean_score >= 0.75 document chars: 200 .. 500000 deterministic stream shuffle seed: 20260906 shuffle buffer:… See the full description on the dataset page: https://huggingface.co/datasets/Efe2898/Prosperity-Family-Alya-CPT-3B-Kumru.tabularn<1K0 likes129 downloads16d agoHugging Face10ProCreations /grug-3b-train grug-3b-train training data for ProCreations/grug-3b. grug think in grug. grug answer in normal english. never other way round. what make this one different old grug model think short always. easy question, short think - good. hard question, short think - BAD. answer come out worse because grug not do the work. this set fix that. every fresh example carry difficulty tier, and tier decide how many word the think get. validator throw away think too short for tier… See the full description on the dataset page: https://huggingface.co/datasets/ProCreations/grug-3b-train.texttext-generation1K<n<10K5 likes110 downloads2mo agoHugging Face11bestdive /details_bestdive__SmolLM3-3B-SFT-Free-Course Smol course SFT evaluation - Kay Zheng Actual full GSM8K test evaluation of bestdive/SmolLM3-3B-SFT-Free-Course, adapter revision 0484e028b494d605a267050a949c9266edadd16b, merged with pinned SmolLM3-3B-Base before evaluation. Full 1319 test examples, zero-shot, original extractive_match: 0.4086429112964367 (stderr 0.013540639733342422). Free Google Colab T4, no paid HF Jobs; cost 0. lighteval 0.11.0, vLLM 0.10.1.1, Transformers 4.57.1, Python 3.12. Dataset-address correction… See the full description on the dataset page: https://huggingface.co/datasets/bestdive/details_bestdive__SmolLM3-3B-SFT-Free-Course.textn<1K0 likes95 downloads13d agoHugging Face12LeeXugar /CodePin-SFT-Qwen3.5-35B-A3B CodePin SFT — Qwen3.5-35B-A3B Teacher Trajectories CodePin SFT contains 6,000 validated code-localization trajectories generated with qwen3.5-35b-a3b. It is designed for pure-text supervised fine-tuning of Qwen/Qwen3.5-0.8B and other tool-calling language models. The tasks come from LeeXugar/SWE-smith-code-search. The source rollout dataset was used only for cleaning, difficulty estimation, and sample selection; rollout messages and rewards were not copied into these SFT… See the full description on the dataset page: https://huggingface.co/datasets/LeeXugar/CodePin-SFT-Qwen3.5-35B-A3B.texttext-generation1K<n<10K0 likes83 downloads1mo agoHugging Face13Hsueh008 /Qwen2.5-3B-jsonltabular100K<n<1M0 likes81 downloads1mo agoHugging Face14visual-memory /ConvAI2-ERNIE-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-ernie-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes80 downloads13d agoHugging Face153B-Group /ConvRetext10K<n<100K1 likes79 downloads2y agoHugging Face16visual-memory /ConvAI2-Qwen-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-qwen-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-Qwen-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-Qwen-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes79 downloads14d agoHugging Face17visual-memory /ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-qwen-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-Qwen-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes75 downloads13d agoHugging Face18ericflo /Llama-3.2-3B-COTtext10K<n<100K0 likes72 downloads2y agoHugging Face19visual-memory /ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-flux-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes72 downloads13d agoHugging Face20visual-memory /ConvAI2-FLUX-original-Qwen3.5-35B-A3B Visual Memory Results: convai2-flux-original This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-FLUX-original-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-FLUX-original-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-FLUX-original-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes68 downloads13d agoHugging Face21JacobMolBio /vibethinker-3b-jlens-traces VibeThinker-3B J Lens Traces The J Lens is a Jacobian lens fitted to WeiboAI/VibeThinker-3B. Pick one of the 18 fitted source layers and a token position, and the lens decodes that residual-stream activation into a ranked list of vocabulary tokens. Following one position across layers shows how the decoded ranking changes on the way to the model's final output. This repository stores saved results from that lens. The companion model repository contains the lens weights. The… See the full description on the dataset page: https://huggingface.co/datasets/JacobMolBio/vibethinker-3b-jlens-traces.tabularn<1K0 likes66 downloads2mo agoHugging Face22visual-memory /ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B Visual Memory Results: convai2-ernie-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/ConvAI2-With-Ids_1k-no-redundancy", "hf_mapping_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/ConvAI2-ERNIE-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes66 downloads13d agoHugging Face23open-llm-leaderboard /NousResearch__Hermes-3-Llama-3.2-3B-detailsgated Dataset Card for Evaluation run of NousResearch/Hermes-3-Llama-3.2-3B Dataset automatically created during the evaluation run of model NousResearch/Hermes-3-Llama-3.2-3B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Hermes-3-Llama-3.2-3B-details.tabular10K<n<100K0 likes61 downloads2y agoHugging Face24open-llm-leaderboard /bigcode__starcoder2-3b-detailsgated Dataset Card for Evaluation run of bigcode/starcoder2-3b Dataset automatically created during the evaluation run of model bigcode/starcoder2-3b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigcode__starcoder2-3b-details.tabular10K<n<100K0 likes60 downloads2y agoHugging Face25visual-memory /PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-flux-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-FLUX-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes58 downloads14d agoHugging Face26visual-memory /PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-ernie-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-ERNIE-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes58 downloads14d agoHugging Face27baohao /Math_Qwen3-30B-A3B-Instruct-2507_SFT-RL_Prefixtext10K<n<100K1 likes57 downloads2mo agoHugging Face28lhchau /llama3.2-3B-sfttext10K<n<100K0 likes56 downloads2mo agoHugging Face29visual-memory /PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B Visual Memory Results: personachat-qwen-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset": "visual-memory/PersonaChat-With-Ids_1k-no-redundancy"… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/PersonaChat-Qwen-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes56 downloads14d agoHugging Face30visual-memory /Synthetic-Persona-Chat-Qwen-enhanced-Qwen3.5-35B-A3B Visual Memory Results: synthetic-persona-chat-qwen-enhanced This dataset contains the scored output of a visual-memory perplexity experiment. Experiment metadata { "experiment": { "model_name": "Qwen/Qwen3.5-35B-A3B", "hf_results_repo": "visual-memory/Synthetic-Persona-Chat-Qwen-enhanced-Qwen3.5-35B-A3B", "results_jsonl": "results/Synthetic-Persona-Chat-Qwen-enhanced-Qwen3.5-35B-A3B.jsonl", "hf_dataset":… See the full description on the dataset page: https://huggingface.co/datasets/visual-memory/Synthetic-Persona-Chat-Qwen-enhanced-Qwen3.5-35B-A3B.tabular1K<n<10K0 likes52 downloads13d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.