datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Chinese-DeepSeek-R1-Distill-data-110k-decontaminated
Decontaminated — Congliu/Chinese-DeepSeek-R1-Distill-data-110k
What this is
A filtered version of Congliu/Chinese-DeepSeek-R1-Distill-data-110k (revision
8520b649430617c2be4490f424d251d09d835ed3) with exact-duplicate rows and rows overlapping standard benchmark test sets
removed. This is a different artifact from the companion contamination report — that one is an
audit of what's wrong; this one is the corpus with those rows actually taken out, ready to train on.… See the full description on the dataset page: https://huggingface.co/datasets/liodon-ai/Chinese-DeepSeek-R1-Distill-data-110k-decontaminated.Chinese-DeepSeek-R1-Distill-data-110k-contamination-report
Contamination Report — Congliu/Chinese-DeepSeek-R1-Distill-data-110k
What this is
A row-level audit of Congliu/Chinese-DeepSeek-R1-Distill-data-110k (revision
8520b649430617c2be4490f424d251d09d835ed3) for exact 13-gram overlap with standard benchmark test sets
(gsm8k, hellaswag, humaneval, mmlu). This is not a filtered copy of the source — it's a new
artifact: a list of which rows overlap which benchmark, plus summary statistics, so anyone
training on the source… See the full description on the dataset page: https://huggingface.co/datasets/liodon-ai/Chinese-DeepSeek-R1-Distill-data-110k-contamination-report.tau2-airline-deepseek-distill
τ²-bench airline · DeepSeek teacher trajectories
Successful multi-turn agent trajectories on τ²-bench's
airline domain, generated by running DeepSeek V4 Flash as the agent through the real τ²-bench
harness — same system prompt, same 14 tool schemas, same dialogue loop, same evaluator.
Used to behavior-clone the RL warm start
yuyu0529nya/qwen2.5-7b-tau2-airline-sft-lora,
which is the step 0 of the tau2_airline verl recipe.
Why these exist
GRPO on τ²-bench-airline… See the full description on the dataset page: https://huggingface.co/datasets/yuyu0529nya/tau2-airline-deepseek-distill.augmented_codealpaca-20k-using-together-ai-deepseek-v1
Dataset Overview
This dataset, named CodeAlpaca-20k, consists of examples that blend coding instructions with outputs and reasoning. Each entry includes structured fields like output, instruction, input, and cot (Chain of Thought). It is particularly designed to train and evaluate AI models that generate code and explanations based on simple programming tasks.
Data Collection and Preparation
Data entries are augmented using the augment_answer function that makes API… See the full description on the dataset page: https://huggingface.co/datasets/eagle0504/augmented_codealpaca-20k-using-together-ai-deepseek-v1.GPT-Deepseek-German
GPT-Deepseek-German QA Dataset
Ein hochwertiger Datensatz mit deutschsprachigen Frage-Antwort-Paaren von DeepSeek und ChatGPT-5 für das Training und Fine-Tuning von Sprachmodellen.
Übersicht
Dieser Datensatz enthält hochqualitative deutschsprachige Konversationsdaten, die aus den KI-Assistenten DeepSeek und ChatGPT-5 generiert wurden. Er eignet sich besonders zum Fine-Tuning von Sprachmodellen im deutschen Sprachraum und für die Entwicklung von Chat-Anwendungen.… See the full description on the dataset page: https://huggingface.co/datasets/Atomic-Ai/GPT-Deepseek-German.deepseek-ai-deepseek-coder-v2-lite-instruct__llm-quality-persona-consistency-mini__019e3b6fdda4
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.quality.persona-consistency-mini (NVIDIA H100 80GB HBM3)
Back to leaderboard
Headline metrics
Metric
Value
Unit
N Samples
5
N Ok
5
Ok Rate
1
Persona Consistency Mean
0.88
Accuracy
0.88
Persona Consistency P50
0.8
Persona Consistency P95
1
Accuracy P05
0.8
Accuracy P50
0.8
Accuracy P95
1
Drift Rate
0.6
Mean Drift Turn
2.6667
TTFT P50
57.5324
ms
Total P50 Ms
2672.8064… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-quality-persona-consistency-mini__019e3b6fdda4.deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short (NVIDIA H100 80GB HBM3)
Back to leaderboard
Headline metrics
Metric
Value
Unit
TTFT P50
74.4268
ms
TTFT P99
521.1922
ms
TPOT P50
21.0371
ms
TPOT P99
23.458
ms
Total P50 Ms
2661.4077
Total P99 Ms
3136.5972
Req Per S Passing
1.0172
Req Per S All
1.1559
Compliance Rate
0.88
Ok Rate
1
Throughput Tok Per S
134.316
Power Avg W
808.3085
Power Peak W
854.62… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca.deepseek-ai-deepseek-coder-v2-lite-instruct__code-generation-humaneval-mini__019e3b6f4f8a
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on code.generation.humaneval-mini (NVIDIA H100 80GB HBM3)
Back to leaderboard
Headline metrics
Metric
Value
Unit
N Samples
5
N Ok
5
Ok Rate
1
Pass At 1
1
Pass At 1 P05
1
Pass At 1 P50
1
Pass At 1 P95
1
Timeout Rate
0
TTFT P50
57.5723
ms
Total P50 Ms
1244.1407
Tokens Out Total
789
Run configuration
Model: deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct @ unknown00
Engine:… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__code-generation-humaneval-mini__019e3b6f4f8a.deepseek-ai-deepseek-coder-v2-lite-instruct__code-generation-mbpp-mini__019e3b6f7d82
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on code.generation.mbpp-mini (NVIDIA H100 80GB HBM3)
Back to leaderboard
Headline metrics
Metric
Value
Unit
N Samples
5
N Ok
5
Ok Rate
1
Pass At 1
1
Pass At 1 P05
1
Pass At 1 P50
1
Pass At 1 P95
1
Timeout Rate
0
TTFT P50
55.0369
ms
Total P50 Ms
1212.8
Tokens Out Total
807
Run configuration
Model: deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct @ unknown00
Engine: vllm… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__code-generation-mbpp-mini__019e3b6f7d82.deepseek-ai-deepseek-coder-v2-lite-instruct__llm-quality-factual-mini__019e3b6f9bdd
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.quality.factual-mini (NVIDIA H100 80GB HBM3)
Back to leaderboard
Headline metrics
Metric
Value
Unit
N Samples
10
N Ok
10
Ok Rate
1
Accuracy
1
Accuracy P05
1
Accuracy P50
1
Accuracy P95
1
TTFT P50
45.0118
ms
Total P50 Ms
359.225
Tokens Out Total
373
Run configuration
Model: deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct @ unknown00
Engine: vllm vunknownQuantization:… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-quality-factual-mini__019e3b6f9bdd.
