datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
open-perfectblend-kimi-linear-regen
Open PerfectBlend Kimi Linear Regen
This dataset regenerates the assistant messages in
mlabonne/open-perfectblend
with Kimi-Linear-48B-A3B-Instruct. It is intended for speculative-decoding
drafter training and related research.
Generation
Source conversation structure and user messages: mlabonne/open-perfectblend
Target model: Kimi-Linear-48B-A3B-Instruct
Temperature: 0.7
Maximum new tokens per assistant turn: 8192
Assistant turns were regenerated sequentially.… See the full description on the dataset page: https://huggingface.co/datasets/Dogacel/open-perfectblend-kimi-linear-regen.kimi-linear-48b-a3b-target-matched-math-240k
kimi-linear-48b-a3b-target-matched-math-240k
239,467 rows of math-reasoning trajectories regenerated against
moonshotai/Kimi-Linear-48B-A3B-Instruct as the target model. Used to train DFlash
speculative-decoding drafters in
la-draftery.
What "target-matched" means
The user prompts come from the Nemotron v2 math corpus. The assistant
completions in this dataset are the target model's own outputs — each
prompt was sent to moonshotai/Kimi-Linear-48B-A3B-Instruct and its… See the full description on the dataset page: https://huggingface.co/datasets/Moonlight556/kimi-linear-48b-a3b-target-matched-math-240k.linear-bench-mini
Agent-Diff: Linear Bench Mini
This dataset is part of the Agent-Diff benchmark, presented in the paper Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation.
Website | GitHub | Paper
Context
The Linear Bench suite runs inside the Agent Diff isolation engine, with its own Postgres schema replaying the Linear GraphQL API. Agents interact via Linear's public surface area to satisfy CRUD-style tasks (create issues… See the full description on the dataset page: https://huggingface.co/datasets/hubertmarek/linear-bench-mini.
