CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sprapp /eagle-grpo-iter19-fp4-enc0 likes886 downloads2mo agoHugging Face02samuelt0207 /Wan2.2-T2V-Activations-FP40 likes784 downloads10mo agoHugging Face03samuelt0207 /LTX-Video-Activations-FP40 likes731 downloads10mo agoHugging Face04sprapp /eagle-sftr2-iter49-fp4-enc0 likes714 downloads2mo agoHugging Face05samuelt0207 /Wan2.2-I2V-Activations-FP40 likes650 downloads10mo agoHugging Face06devpjai /tg-glm-fp4-pool0 likes179 downloads14d agoHugging Face07JacobChang /GLM5.2-FP4-MI355X-profiles0 likes26 downloads2mo agoHugging Face08rand0nmr /fp4_1.3Bvideon<1K0 likes17 downloads8mo agoHugging Face09daniehua /kimik25-fp4-vllm-isl8192osl1024conc128tabularn<1K0 likes17 downloads4mo agoHugging Face10chengfu0118 /DeepSeek-R1-FP4_1757573629_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757573629_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 61.11% ± 0.48% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 61.11% 121 198 2 60.10% 119 198 3 62.12% 123 198 tabularn<1K0 likes13 downloads1y agoHugging Face11open-llm-leaderboard /ehristoforu__fp4-14b-v1-fix-detailsgated Dataset Card for Evaluation run of ehristoforu/fp4-14b-v1-fix Dataset automatically created during the evaluation run of model ehristoforu/fp4-14b-v1-fix The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ehristoforu__fp4-14b-v1-fix-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face12chengfu0118 /DeepSeek-R1-FP4_1757550665_eval_d91d chengfu0118/DeepSeek-R1-FP4_1757550665_eval_d91d Precomputed model outputs for evaluation. Evaluation Results MATH500 Accuracy: 88.60% Accuracy Questions Solved Total Questions 88.60% 443 500 tabularn<1K0 likes12 downloads1y agoHugging Face13chengfu0118 /DeepSeek-R1-FP4_1757557889_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757557889_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 61.28% ± 0.77% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 63.13% 125 198 2 60.61% 120 198 3 60.10% 119 198 tabularn<1K0 likes10 downloads1y agoHugging Face14daniehua /dsr1-fp4-sgl-isl8192osl1024tabularn<1K0 likes9 downloads7mo agoHugging Face15daniehua /dsr1-fp4-sgl-isl1024osl8192tabularn<1K0 likes8 downloads7mo agoHugging Face16daniehua /kimik25-fp4-vllm-isl8192osl1024conc32tabularn<1K0 likes8 downloads4mo agoHugging Face17chengfu0118 /DeepSeek-R1-FP4_1757548772_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757548772_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 59.26% ± 1.07% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 61.62% 122 198 2 57.07% 113 198 3 59.09% 117 198 tabularn<1K0 likes7 downloads1y agoHugging Face18chengfu0118 /DeepSeek-R1-FP4_1757540347_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757540347_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 16.67% ± 0.63% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 17.17% 34 198 2 17.68% 35 198 3 15.15% 30 198 tabularn<1K0 likes7 downloads1y agoHugging Face19chengfu0118 /DeepSeek-R1-FP4_1757570920_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757570920_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 59.93% ± 0.77% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 60.61% 120 198 2 61.11% 121 198 3 58.08% 115 198 tabularn<1K0 likes7 downloads1y agoHugging Face20daniehua /gptoss-fp4-vllm-isl1024osl1024tabularn<1K0 likes7 downloads8mo agoHugging Face21daniehua /gptoss-fp4-vllm-isl8192osl1024tabularn<1K0 likes7 downloads10mo agoHugging Face22daniehua /dsr1-fp4-sgl-isl8192osl1024conc4tabularn<1K0 likes7 downloads4mo agoHugging Face23open-llm-leaderboard /ehristoforu__fp4-14b-it-v1-detailsgated Dataset Card for Evaluation run of ehristoforu/fp4-14b-it-v1 Dataset automatically created during the evaluation run of model ehristoforu/fp4-14b-it-v1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ehristoforu__fp4-14b-it-v1-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face24chengfu0118 /DeepSeek-R1-FP4_1757536719_eval_2870 chengfu0118/DeepSeek-R1-FP4_1757536719_eval_2870 Precomputed model outputs for evaluation. Evaluation Results AIME24 Average Accuracy: 68.33% ± 1.18% Number of Runs: 10 Run Accuracy Questions Solved Total Questions 1 70.00% 21 30 2 63.33% 19 30 3 66.67% 20 30 4 63.33% 19 30 5 70.00% 21 30 6 70.00% 21 30 7 66.67% 20 30 8 70.00% 21 30 9 66.67% 20 30 10 76.67% 23 30 tabularn<1K0 likes6 downloads1y agoHugging Face25chengfu0118 /DeepSeek-R1-FP4_1757537053_eval_2870 chengfu0118/DeepSeek-R1-FP4_1757537053_eval_2870 Precomputed model outputs for evaluation. Evaluation Results AIME24 Average Accuracy: 65.00% ± 2.27% Number of Runs: 10 Run Accuracy Questions Solved Total Questions 1 73.33% 22 30 2 70.00% 21 30 3 70.00% 21 30 4 56.67% 17 30 5 56.67% 17 30 6 73.33% 22 30 7 66.67% 20 30 8 70.00% 21 30 9 53.33% 16 30 10 60.00% 18 30 tabularn<1K0 likes6 downloads1y agoHugging Face26chengfu0118 /DeepSeek-R1-FP4_1757541727_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757541727_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 60.10% ± 1.33% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 63.13% 125 198 2 59.60% 118 198 3 57.58% 114 198 tabularn<1K0 likes6 downloads1y agoHugging Face27chengfu0118 /DeepSeek-R1-FP4_1757547837_eval_d91d chengfu0118/DeepSeek-R1-FP4_1757547837_eval_d91d Precomputed model outputs for evaluation. Evaluation Results MATH500 Accuracy: 88.20% Accuracy Questions Solved Total Questions 88.20% 441 500 tabularn<1K0 likes6 downloads1y agoHugging Face28chengfu0118 /DeepSeek-R1-FP4_1757561992_eval_f912 chengfu0118/DeepSeek-R1-FP4_1757561992_eval_f912 Precomputed model outputs for evaluation. Evaluation Results GPQADiamond Average Accuracy: 60.44% ± 1.91% Number of Runs: 3 Run Accuracy Questions Solved Total Questions 1 64.65% 128 198 2 56.57% 112 198 3 60.10% 119 198 tabularn<1K0 likes6 downloads1y agoHugging Face29chengfu0118 /DeepSeek-R1-FP4_1757543283_eval_7c1d chengfu0118/DeepSeek-R1-FP4_1757543283_eval_7c1d Precomputed model outputs for evaluation. Evaluation Results Summary Metric AIME24 MATH500 GPQADiamond Accuracy 0.7 15.0 7.4 AIME24 Average Accuracy: 0.67% ± 0.63% Number of Runs: 10 Run Accuracy Questions Solved Total Questions 1 0.00% 0 30 2 0.00% 0 30 3 0.00% 0 30 4 0.00% 0 30 5 6.67% 2 30 6 0.00% 0 30 7 0.00% 0 30 8 0.00% 0 30 9 0.00% 0 30 10 0.00% 0… See the full description on the dataset page: https://huggingface.co/datasets/chengfu0118/DeepSeek-R1-FP4_1757543283_eval_7c1d.tabular1K<n<10K0 likes6 downloads1y agoHugging Face30daniehua /dsr1-fp4-sgl-isl8192osl1024conc32tabularn<1K0 likes6 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.