DeepSeekV3
DeepSeek-V3-Infinity-Instruct-0625
DeepSeek-V3-Infinity-Instruct-0625
Dataset Description
This dataset is part of the LK-Speculators collection for speculative decoding research. It contains 660K prompt-response pairs designed for training draft models that are used alongside DeepSeek-V3-0324 as the target model. The dataset was created by generating responses to the prompts from Infinity-Instruct-0625 with deepseek-ai/DeepSeek-V3-0324 at temperature=1.
For more details on the training methodology and… See the full description on the dataset page: https://huggingface.co/datasets/nebius/DeepSeek-V3-Infinity-Instruct-0625.deepseek-v3.2-speciale-openr1-math-3kInspired by @OpenR1
The questions for this dataset were all sourced from the first 3.3k prompts in open-r1/OpenR1-Math-220k
Dataset Stats (provided by OpenRouter):
Cost: $ 21.1 (USD)
Tokens (input + output): 52.3 M
deepseek-v3.2-speciale-OpenCodeReasoning-3kThe questions for this dataset were all sourced from the first 3k prompts in nvidia/OpenCodeReasoning
Dataset Stats (provided by OpenRouter):
Cost: $ 19.2 (USD)
Tokens (input + output): 47 M
deepseek-v3-10kThe first 10K elements of The Pile, useful for debugging models trained on it. See the HuggingFace page for the full Pile for more info. Inspired by stas' great resource doing the same for OpenWebText
deepseekv3-ultrafeedback-armorm-dpochemistry-reasoning-phi4-vs-deepseekv3
