datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DrafterBench-fixed-trajectories
AgentSuite/DrafterBench-fixed-trajectories
Per-model agent trajectory data for DrafterBench-fixed (public release).
Models: 30
Tasks per model: 1,920
One file per model: {model}.jsonl, one JSON object per line.
Fields: model_path, user_model_path, benchmark_name, task_name, sampling_params, user_sampling_params, messages, eval_result, meta.
sampling_params reflect each benchmark's own implementation; values the benchmark leaves unset are recorded as null (provider default).… See the full description on the dataset page: https://huggingface.co/datasets/AgentSuite/DrafterBench-fixed-trajectories.DrafterBench
Dataset Card for DrafterBench
Dataset Description
DrafterBench is a large-scale toolkit focused on evaluating the proficiency of Large Language Models (LLMs) in automating Civil Engineering tasks.
This dataset hosts a task suite summarized across 20 real-world projects, encompassing a total of 1920 tasks.
It replicates the complexity of real-world engineering tasks and provides a technical platform to test the four key capabilities of LLMs:
Structured data understanding… See the full description on the dataset page: https://huggingface.co/datasets/Eason666/DrafterBench.turkish-flow-drafter-prompts
Turkish prompts for Chained-Flow drafter training
Chat-templated Turkish prompts used to train and evaluate the Turkish
Flow-Drafter
checkpoints for Qwen/Qwen3.5-4B / 9B / 27B.
Prompts only — no completions. A drafter is trained on the target model's own hidden states, so
continuations are generated locally by running the target over these prompts. Nothing here is a
model output.
split
rows
what it is
v1/
29,100 train + 300 holdout
the mixture the released Turkish… See the full description on the dataset page: https://huggingface.co/datasets/selimaktas/turkish-flow-drafter-prompts.turkish-flow-drafter-prompts
GitHub repo ·
Technical blog ·
Model collection
Turkish prompts for Chained-Flow drafter training
Chat-templated Turkish prompts used to train and evaluate the Turkish
Flow-Drafter
checkpoints for Qwen/Qwen3.5-4B / 9B / 27B.
Prompts only — no completions. A drafter is trained on the target model's own hidden states, so
continuations are generated locally by running the target over these prompts. Nothing here is a
model output.
split
rows
what it is
v1/… See the full description on the dataset page: https://huggingface.co/datasets/ytu-ce-cosmos/turkish-flow-drafter-prompts.english-flow-drafter-prompts
English prompts for Chained-Flow drafter training
Chat-templated English prompts used to train the English
Flow-Drafter
checkpoints for Qwen/Qwen3.5-4B / 9B / 27B.
Prompts only — no completions. A drafter is trained on the target model's own hidden states, so
continuations are generated locally by running the target over these prompts. Nothing here is a
model output.
split
rows
prompt tokens
what it is
v1/
19,672 train + 500 holdout
1,518,620
the original mixture… See the full description on the dataset page: https://huggingface.co/datasets/selimaktas/english-flow-drafter-prompts.claim_drafter
Claim Drafter — datasets
Training and evaluation data for vishwr/claim_drafter,
a LoRA adapter that drafts US patent claims from a plain-English invention disclosure.
Two datasets are included:
sft — 9,662 train / 1,314 validation examples in conversational (messages)
format. Targets are as-GRANTED claims fetched per patent number (not the
as-filed claims that ship with HUPD), and prompts are stripped of the patent
summary so the model must draft rather than reformat.
dpo — 5… See the full description on the dataset page: https://huggingface.co/datasets/vishwr/claim_drafter.english-flow-drafter-prompts
GitHub repo ·
Technical blog ·
Model collection
English prompts for Chained-Flow drafter training
Chat-templated English prompts used to train the English
Flow-Drafter
checkpoints for Qwen/Qwen3.5-4B / 9B / 27B.
Prompts only — no completions. A drafter is trained on the target model's own hidden states, so
continuations are generated locally by running the target over these prompts. Nothing here is a
model output.
split
rows
prompt tokens
what it is
v1/… See the full description on the dataset page: https://huggingface.co/datasets/ytu-ce-cosmos/english-flow-drafter-prompts.DrafterBench
Dataset Card for DrafterBench
DrafterBench
DrafterBench is a large-scale toolkit focused on evaluating the proficiency of Large Language Models (LLMs) in automating Civil Engineering tasks.
The dataset contains tasks derived from real-world engineering drawing revision processes.
This dataset is released for anonymous review.
Code: https://github.com/anonymous733882/DrafterBench
This dataset hosts a task suite summarized across 20 real-world projects, encompassing a total… See the full description on the dataset page: https://huggingface.co/datasets/anonymous733882/DrafterBench.bonsai2-drafter-eval
Bonsai 2 drafter evaluation corpora
Prompts and greedy responses from PrismML's
Ternary-Bonsai-2-27B, recorded
as token ids together with the per-round accepted lengths of the speculative decoding loop that
produced them. The data exists to evaluate and train DFlash 2 drafters against this one target.
Target
prism-ml/Ternary-Bonsai-2-27B-mlx-2bit (MLX pack), greedy, temperature 0
Loop
mlx-dspark 0.18.0 DFlash 2 loop, stock drafter z-lab/Qwen3.8-27B-DFlash2, draft… See the full description on the dataset page: https://huggingface.co/datasets/Schiltmans/bonsai2-drafter-eval.
