datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
open-perfectblend-kimi-linear-regen
Open PerfectBlend Kimi Linear Regen
This dataset regenerates the assistant messages in
mlabonne/open-perfectblend
with Kimi-Linear-48B-A3B-Instruct. It is intended for speculative-decoding
drafter training and related research.
Generation
Source conversation structure and user messages: mlabonne/open-perfectblend
Target model: Kimi-Linear-48B-A3B-Instruct
Temperature: 0.7
Maximum new tokens per assistant turn: 8192
Assistant turns were regenerated sequentially.… See the full description on the dataset page: https://huggingface.co/datasets/Dogacel/open-perfectblend-kimi-linear-regen.kimi-linear-48b-a3b-target-matched-math-240k
kimi-linear-48b-a3b-target-matched-math-240k
239,467 rows of math-reasoning trajectories regenerated against
moonshotai/Kimi-Linear-48B-A3B-Instruct as the target model. Used to train DFlash
speculative-decoding drafters in
la-draftery.
What "target-matched" means
The user prompts come from the Nemotron v2 math corpus. The assistant
completions in this dataset are the target model's own outputs — each
prompt was sent to moonshotai/Kimi-Linear-48B-A3B-Instruct and its… See the full description on the dataset page: https://huggingface.co/datasets/Moonlight556/kimi-linear-48b-a3b-target-matched-math-240k.
