CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jonathanyin /aime_1983_2023_deepseek-r1_traces_16384tabularn<1K0 likes3.2k downloads1y agoHugging Face02Rock23210 /AIME_Deepseek_Cleantextn<1K0 likes1.2k downloads2y agoHugging Face03hbXNov /numina_amc_aime_deepseek_r1_responsestextn<1K0 likes1.2k downloads2y agoHugging Face04fireworks-ai /logiqa-deepseek-v3text1K<n<10K0 likes1k downloads2y agoHugging Face05deepseek-ai /DeepSeek-Prover-V1 Evaluation Results | Model & Dataset Downloads | License | Contact Paper Link👁️ DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data 1. Introduction Proof assistants like Lean have revolutionized mathematical proof verification, ensuring high accuracy and reliability. Although large language models (LLMs) show promise in… See the full description on the dataset page: https://huggingface.co/datasets/deepseek-ai/DeepSeek-Prover-V1.text10K<n<100K74 likes895 downloads2y agoHugging Face06jonathanyin /aime_1983_2023_deepseek-r1_traces_32768tabularn<1K0 likes809 downloads1y agoHugging Face07deepseek-ai /DeepSeek-ProverBench 1. Introduction We introduce DeepSeek-Prover-V2, an open-source large language model designed for formal theorem proving in Lean 4, with initialization data collected through a recursive theorem proving pipeline powered by DeepSeek-V3. The cold-start training procedure begins by prompting DeepSeek-V3 to decompose complex problems into a series of subgoals. The proofs of resolved subgoals… See the full description on the dataset page: https://huggingface.co/datasets/deepseek-ai/DeepSeek-ProverBench.textn<1K47 likes654 downloads1y agoHugging Face08jonathanyin /aime_1983_2023_deepseek-r1-distill-qwen-14b_traces_32768tabularn<1K0 likes474 downloads1y agoHugging Face09jonathanyin /aime_1983_2023_deepseek-r1-distill-qwen-7b_traces_32768tabularn<1K0 likes444 downloads1y agoHugging Face10jonathanyin /aime_1983_2023_deepseek-r1-distill-qwen-1.5b_traces_32768tabularn<1K0 likes401 downloads1y agoHugging Face11NodeLinker /deepseek-ai-Thinking-with-Visual-Primitives-deleted-repo Thinking with Visual Primitives English | 简体中文 News 2026.04.30: We have released the technical report detailing our approach. In the near future, we plan to make the in-house benchmarks and a subset of our cold-start data publicly available. The model weights will be integrated into our foundation model and released in the future. 1. Introduction While recent Multimodal Large Language Models (MLLMs) have made strides in… See the full description on the dataset page: https://huggingface.co/datasets/NodeLinker/deepseek-ai-Thinking-with-Visual-Primitives-deleted-repo.40 likes340 downloads5mo agoHugging Face12guanning-ai /DeepSeek-1.5B_mmlu-pro_16384_train0test8tabular100K<n<1M0 likes252 downloads1y agoHugging Face13arcee-ai /DeepSeek-MixedModeReasoning-Logits-Packed-16384sequence_length: 16384 dataset: train_dataset: repo_id: arcee-ai/DeepSeek-MixedModeReasoning-Logits-Packed-16384 split: train prepacked: true teacher: kind: dataset legacy_logit_compression: exact_k: 32 invert_polynomial: true k: 32 polynomial_degree: 0 term_dtype: float32 vocab_size: 129280 with_sqrt_term: false 100K<n<1M8 likes230 downloads10mo agoHugging Face14OALL /details_deepseek-ai__DeepSeek-R1-Distill-Qwen-14B_v2 Dataset Card for Evaluation run of deepseek-ai/DeepSeek-R1-Distill-Qwen-14B Dataset automatically created during the evaluation run of model deepseek-ai/DeepSeek-R1-Distill-Qwen-14B. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_deepseek-ai__DeepSeek-R1-Distill-Qwen-14B_v2.text100K<n<1M0 likes212 downloads2y agoHugging Face15jonathanyin /aime_1983_2023_deepseek-r1-distill-qwen-7b_traces_16384tabularn<1K0 likes202 downloads1y agoHugging Face16OALL /details_deepseek-ai__DeepSeek-R1-Distill-Qwen-32B_v2 Dataset Card for Evaluation run of deepseek-ai/DeepSeek-R1-Distill-Qwen-32B Dataset automatically created during the evaluation run of model deepseek-ai/DeepSeek-R1-Distill-Qwen-32B. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_deepseek-ai__DeepSeek-R1-Distill-Qwen-32B_v2.text100K<n<1M0 likes194 downloads2y agoHugging Face17guruswami-ai /deepseek-v4-flash-0731-m3-ultra DeepSeek-V4-Flash-0731 on M3 Ultra 512 GB — benchmark dataset Independent performance characterization of Vontra/DeepSeek-V4-Flash-0731-MXFP4-MLX on a single Mac Studio M3 Ultra (80-core GPU, 512 GB unified memory). Engine: omlx 0.5.7 · OS: macOS 26.6 (25G72) · MLX: 0.32.0 Recommended configuration omlx serve --model-dir /opt/models --port 8033 \ --hot-cache-max-size 256GB --initial-cache-blocks 512 // ~/.omlx/model_settings.json {"version": 1, "models":… See the full description on the dataset page: https://huggingface.co/datasets/guruswami-ai/deepseek-v4-flash-0731-m3-ultra.imagen<1K0 likes187 downloads2mo agoHugging Face18open-llm-leaderboard-old /details_AIGym__deepseek-coder-6.7b-chat Dataset Card for Evaluation run of AIGym/deepseek-coder-6.7b-chat Dataset automatically created during the evaluation run of model AIGym/deepseek-coder-6.7b-chat on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_AIGym__deepseek-coder-6.7b-chat.0 likes170 downloads3y agoHugging Face19CreitinGameplays /DeepSeek-R1-Distill-Qwen-32B_NUMINA_train_amc_aime-llama3.1tabular1K<n<10K0 likes167 downloads2y agoHugging Face20hbXNov /numina_amc_aime_in_depth_deepseek_r1_questionstextn<1K0 likes156 downloads2y agoHugging Face21open-llm-leaderboard-old /details_deepseek-ai__deepseek-coder-1.3b-instruct Dataset Card for Evaluation run of deepseek-ai/deepseek-coder-1.3b-instruct Dataset Summary Dataset automatically created during the evaluation run of model deepseek-ai/deepseek-coder-1.3b-instruct on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_deepseek-ai__deepseek-coder-1.3b-instruct.0 likes148 downloads3y agoHugging Face22open-llm-leaderboard-old /details_deepseek-ai__deepseek-math-7b-instruct Dataset Card for Evaluation run of deepseek-ai/deepseek-math-7b-instruct Dataset automatically created during the evaluation run of model deepseek-ai/deepseek-math-7b-instruct on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_deepseek-ai__deepseek-math-7b-instruct.0 likes140 downloads3y agoHugging Face23open-llm-leaderboard-old /details_deepseek-ai__deepseek-coder-6.7b-base Dataset Card for Evaluation run of deepseek-ai/deepseek-coder-6.7b-base Dataset automatically created during the evaluation run of model deepseek-ai/deepseek-coder-6.7b-base on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_deepseek-ai__deepseek-coder-6.7b-base.0 likes138 downloads2y agoHugging Face24OALL /details_deepseek-ai__DeepSeek-R1-Distill-Llama-70B Dataset Card for Evaluation run of deepseek-ai/DeepSeek-R1-Distill-Llama-70B Dataset automatically created during the evaluation run of model deepseek-ai/DeepSeek-R1-Distill-Llama-70B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_deepseek-ai__DeepSeek-R1-Distill-Llama-70B.tabular100K<n<1M0 likes137 downloads2y agoHugging Face25open-llm-leaderboard-old /details_deepseek-ai__deepseek-llm-67b-chat Dataset Card for Evaluation run of deepseek-ai/deepseek-llm-67b-chat Dataset automatically created during the evaluation run of model deepseek-ai/deepseek-llm-67b-chat on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_deepseek-ai__deepseek-llm-67b-chat.0 likes131 downloads3y agoHugging Face26guanning-ai /DeepSeek-1.5B_math_32768_train8test64tabular100K<n<1M0 likes127 downloads1y agoHugging Face27OALL /details_huihui-ai__DeepSeek-R1-Distill-Qwen-32B-abliterated Dataset Card for Evaluation run of huihui-ai/DeepSeek-R1-Distill-Qwen-32B-abliterated Dataset automatically created during the evaluation run of model huihui-ai/DeepSeek-R1-Distill-Qwen-32B-abliterated. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_huihui-ai__DeepSeek-R1-Distill-Qwen-32B-abliterated.tabular100K<n<1M0 likes124 downloads2y agoHugging Face28jonathanyin /aime_1983_2023_deepseek-r1-distill-qwen-14b_traces_16384tabularn<1K0 likes124 downloads1y agoHugging Face29vectorzhou /AIME_2024_DeepSeek_R1_0528_Temp_1.0_L_16384Responses of deepseek-ai/DeepSeek-R1-0528 for AIME 2024 (original dataset: Maxwell-Jia/AIME_2024). Generation temperature is set to 1.0 and maximum token is set to 16384. textn<1K0 likes123 downloads1y agoHugging Face30guanning-ai /DeepSeek-1.5B_dapo2k_16384_train32test0tabular10K<n<100K0 likes118 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.