CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01joshycodes /sorrel-T-qwen3-14b-base-seed0-documentstext100K<n<1M0 likes191 downloads5d agoHugging Face02twinkle-ai /ministral-14b-eval-logs-and-scorestabular100K<n<1M0 likes159 downloads7mo agoHugging Face03dongboklee /gPRM-14B-test_qwen Reward of test_qwen split extracted by gPRM-14B: gPRM-14B-test_qwen Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/gPRM-14B-test_qwen") # Load specific domain law_dataset = load_dataset("dongboklee/gPRM-14B-test_qwen", split="law") text1K<n<10K0 likes113 downloads1y agoHugging Face04nyu-dice-lab /lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_14B-private Dataset Card for Evaluation run of TomGrc/FusionNet_7Bx2_MoE_14B Dataset automatically created during the evaluation run of model TomGrc/FusionNet_7Bx2_MoE_14B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_14B-private.tabular100K<n<1M0 likes112 downloads2y agoHugging Face05dongboklee /dORM-14B-test_gemma Reward of test_gemma split extracted by dORM-14B: dORM-14B-test_gemma Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_gemma") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_gemma", split="law") text1K<n<10K0 likes101 downloads1y agoHugging Face06dongboklee /gPRM-14B-test_llama Reward of test_llama split extracted by gPRM-14B: gPRM-14B-test_llama Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/gPRM-14B-test_llama") # Load specific domain law_dataset = load_dataset("dongboklee/gPRM-14B-test_llama", split="law") text1K<n<10K0 likes96 downloads1y agoHugging Face07dongboklee /dORM-14B-test_qwen Reward of test_qwen split extracted by dORM-14B: dORM-14B-test_qwen Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_qwen") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_qwen", split="law") text1K<n<10K0 likes93 downloads1y agoHugging Face08dongboklee /gORM-14B-test_llama Reward of test_llama split extracted by dORM-14B: dORM-14B-test_llama Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_llama") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_llama", split="law") text1K<n<10K0 likes88 downloads1y agoHugging Face09open-llm-leaderboard /suayptalha__Lix-14B-v0.1-detailsgated Dataset Card for Evaluation run of suayptalha/Lix-14B-v0.1 Dataset automatically created during the evaluation run of model suayptalha/Lix-14B-v0.1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/suayptalha__Lix-14B-v0.1-details.tabular10K<n<100K0 likes76 downloads2y agoHugging Face10dongboklee /dORM-14B-test_llama Reward of test_llama split extracted by dORM-14B: dORM-14B-test_llama Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_llama") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_llama", split="law") text1K<n<10K0 likes74 downloads1y agoHugging Face11dongboklee /dPRM-14B-test Reward of test split extracted by dPRM-14B: dPRM-14B-test Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dPRM-14B-test") # Load specific domain law_dataset = load_dataset("dongboklee/dPRM-14B-test", split="law") text1K<n<10K0 likes74 downloads1y agoHugging Face12open-llm-leaderboard /JungZoona__T3Q-qwen2.5-14b-v1.0-e3-detailsgated Dataset Card for Evaluation run of JungZoona/T3Q-qwen2.5-14b-v1.0-e3 Dataset automatically created during the evaluation run of model JungZoona/T3Q-qwen2.5-14b-v1.0-e3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/JungZoona__T3Q-qwen2.5-14b-v1.0-e3-details.tabular10K<n<100K0 likes72 downloads2y agoHugging Face13open-llm-leaderboard /prithivMLmods__Gaea-Opus-14B-Exp-detailsgated Dataset Card for Evaluation run of prithivMLmods/Gaea-Opus-14B-Exp Dataset automatically created during the evaluation run of model prithivMLmods/Gaea-Opus-14B-Exp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Gaea-Opus-14B-Exp-details.tabular10K<n<100K0 likes70 downloads2y agoHugging Face14ApacheOne /Info_Wan_Video_2.2_T2V-A14B Model Index by Creator 423748 Page Model Base Model Full Model Page Archive Link wan2.2,t2v,low,zzzyixuan. Wan Video 2.2 T2V-A14B View View Version Links Model Version Base Model Version Link wan2.2,t2v,low,zzzyixuan. v1.0 Wan Video 2.2 T2V-A14B View Aaron_PP Page Model Base Model Full Model Page Archive Link NSFW WAN 2.2 T2V Bunny girl, red patent leather tights, black high stockings, red high heels Wan Video 2.2… See the full description on the dataset page: https://huggingface.co/datasets/ApacheOne/Info_Wan_Video_2.2_T2V-A14B.textn<1K7 likes66 downloads1y agoHugging Face15dongboklee /dORM-14B-test Reward of test split extracted by dORM-14B: dORM-14B-test Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test", split="law") text1K<n<10K0 likes63 downloads1y agoHugging Face16open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8-details.tabular10K<n<100K0 likes58 downloads2y agoHugging Face17open-llm-leaderboard /Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.9-detailsgated Dataset Card for Evaluation run of Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.9 Dataset automatically created during the evaluation run of model Lunzima/NQLSG-Qwen2.5-14B-MegaFusion-v8.9 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lunzima__NQLSG-Qwen2.5-14B-MegaFusion-v8.9-details.tabular10K<n<100K0 likes58 downloads2y agoHugging Face18open-llm-leaderboard /CultriX__SeQwence-14B-detailsgated Dataset Card for Evaluation run of CultriX/SeQwence-14B Dataset automatically created during the evaluation run of model CultriX/SeQwence-14B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__SeQwence-14B-details.tabular10K<n<100K0 likes57 downloads2y agoHugging Face19open-llm-leaderboard /mrm8488__phi-4-14B-grpo-gsm8k-3e-detailsgated Dataset Card for Evaluation run of mrm8488/phi-4-14B-grpo-gsm8k-3e Dataset automatically created during the evaluation run of model mrm8488/phi-4-14B-grpo-gsm8k-3e The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mrm8488__phi-4-14B-grpo-gsm8k-3e-details.tabular10K<n<100K0 likes55 downloads2y agoHugging Face20open-llm-leaderboard /YOYO-AI__Qwen2.5-14B-YOYO-V4-detailsgated Dataset Card for Evaluation run of YOYO-AI/Qwen2.5-14B-YOYO-V4 Dataset automatically created during the evaluation run of model YOYO-AI/Qwen2.5-14B-YOYO-V4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/YOYO-AI__Qwen2.5-14B-YOYO-V4-details.tabular10K<n<100K0 likes55 downloads2y agoHugging Face21shi3z /ja_conv_wikipedia_orion14B_100K Abstruct This is a multi-turn conversation dataset generated from the Japanese Wikipedia dataset using Orion14B-Chat. Commercial use is possible, but the license is complicated, so please read it carefully before using it. I generated V100x4 on 200 machines in about half a week. License 【Orion-14B Series】 Models Community License Agreement https://huggingface.co/OrionStarAI/Orion-14B-Chat/blob/main/ModelsCommunityLicenseAgreement Computing ABCI… See the full description on the dataset page: https://huggingface.co/datasets/shi3z/ja_conv_wikipedia_orion14B_100K.text100K<n<1M3 likes53 downloads3y agoHugging Face22dongboklee /gORM-14B-test Reward of test split extracted by dORM-14B: dORM-14B-test Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test", split="law") text1K<n<10K0 likes51 downloads1y agoHugging Face23open-llm-leaderboard /sometimesanotion__Qwenvergence-14B-v12-Prose-DS-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.tabular10K<n<100K0 likes50 downloads2y agoHugging Face24Minsang /TSD-KD-Qwen2.5-14B-Instruct-Gen TSD-KD-Qwen2.5-1.5B-Instruct-Gen This dataset contains student-generated examples used for Token-Selective Dual Knowledge Distillation (TSD-KD), introduced in our ICLR 2026 paper: "Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation" Paper: https://arxiv.org/abs/2603.13260 Github: https://github.com/kmswin1/TSD-KD Dataset Description This dataset contains teacher-generated instruction-response examples from… See the full description on the dataset page: https://huggingface.co/datasets/Minsang/TSD-KD-Qwen2.5-14B-Instruct-Gen.texttext-generation10K<n<100K0 likes50 downloads5mo agoHugging Face25open-llm-leaderboard /HeraiHench__Phi-4-slerp-ReasoningRP-14B-detailsgated Dataset Card for Evaluation run of HeraiHench/Phi-4-slerp-ReasoningRP-14B Dataset automatically created during the evaluation run of model HeraiHench/Phi-4-slerp-ReasoningRP-14B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HeraiHench__Phi-4-slerp-ReasoningRP-14B-details.tabular10K<n<100K0 likes49 downloads2y agoHugging Face26open-llm-leaderboard /YOYO-AI__Qwen2.5-14B-it-restore-detailsgated Dataset Card for Evaluation run of YOYO-AI/Qwen2.5-14B-it-restore Dataset automatically created during the evaluation run of model YOYO-AI/Qwen2.5-14B-it-restore The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/YOYO-AI__Qwen2.5-14B-it-restore-details.tabular10K<n<100K0 likes48 downloads2y agoHugging Face27dongboklee /gORM-14B-test_smollm Reward of test_smollm split extracted by dORM-14B: dORM-14B-test_smollm Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_smollm") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_smollm", split="law") text1K<n<10K0 likes48 downloads1y agoHugging Face28dongboklee /gORM-14B-test_qwen Reward of test_qwen split extracted by dORM-14B: dORM-14B-test_qwen Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/dORM-14B-test_qwen") # Load specific domain law_dataset = load_dataset("dongboklee/dORM-14B-test_qwen", split="law") text1K<n<10K0 likes46 downloads1y agoHugging Face29dongboklee /gPRM-14B-test_smollm Reward of test_smollm split extracted by gPRM-14B: gPRM-14B-test_smollm Usage from datasets import load_dataset # Load entire dataset dataset = load_dataset("dongboklee/gPRM-14B-test_smollm") # Load specific domain law_dataset = load_dataset("dongboklee/gPRM-14B-test_smollm", split="law") text1K<n<10K0 likes46 downloads1y agoHugging Face30open-llm-leaderboard /sometimesanotion__Qwen-2.5-14B-Virmarckeoso-detailsgated Dataset Card for Evaluation run of sometimesanotion/Qwen-2.5-14B-Virmarckeoso Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-2.5-14B-Virmarckeoso The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details.tabular10K<n<100K0 likes45 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.