dflash
Datasets
All datasets matching “dflash”osmqwopus-dflash-article-assetsqwopus-dflash-swe20-runtime-results
Qwopus / DFlash SWE20 Runtime Results
Local RTX 3090 Ti benchmark artifacts for 20 long SWE-bench Lite prompts. The run compares Qwopus 3.6 GGUF variants, llama.cpp MTP speculative decoding, QuinsZouls, and DFlash DDTree configurations at 64K context with q8/q8 KV unless noted.
The quality score is a reproducible proxy rubric over gold-patch signals, not official SWE-bench pass/fail. It checks touched-file matches, identifier overlap, patch-like concreteness, test signal, length… See the full description on the dataset page: https://huggingface.co/datasets/jakeatx/qwopus-dflash-swe20-runtime-results.dflash-code-multilingual-teacher-responses-qwen235b
Code + Multilingual Teacher Responses (Qwen3-235B-A22B-Instruct-2507)
This repo now contains 302,800 total samples across the main blended
data.jsonl / .parquet file plus a second Nemotron-only file
(nemotron_code_teacher_responses.jsonl / .parquet). All responses were
generated by Qwen3-235B-A22B-Instruct-2507 in non-thinking mode
(enable_thinking=false) to match downstream speculator training and eval.
Built in two batches: an initial 59,506-row batch (50K code + 9.5K… See the full description on the dataset page: https://huggingface.co/datasets/inference-optimization/dflash-code-multilingual-teacher-responses-qwen235b.MoS-DFlash-Evidence
MoS-DFlash aggregate experiment evidence
This dataset repository contains aggregate, reviewer-facing evidence for the
MoS-DFlash experiments. It does not contain prompts, per-prompt generations,
training data, credentials, internal paths, or raw training logs.
B5 Qwen3-4B fixed-budget replication
releases/b5-qwen3-4b-fixed-budget/ contains:
the frozen result summary;
the matched-training-volume aggregate trajectory;
plot-ready aggregate trajectories;
run… See the full description on the dataset page: https://huggingface.co/datasets/ryan-0608/MoS-DFlash-Evidence.qwen3-4b-dflash-official100k-prepared
Qwen3-4B DFlash Official-100K Prepared Dataset
This is the prepared datasets.load_from_disk() artifact used by the
Qwen3-4B official-route DFlash recipe:
100,000 examples
columns: input_ids, loss_mask, seq_len
max training sequence length used by the recipe: 3072
route: Qwen3 no-thinking / enable_thinking=false
Use it with:
from huggingface_hub import snapshot_download
from datasets import load_from_disk
path = snapshot_download(… See the full description on the dataset page: https://huggingface.co/datasets/jiamingshan/qwen3-4b-dflash-official100k-prepared.qwen3.8-27b-uncensored-dflash2-m4-pro-benchmark
Qwen3.8-27B (MLX 4-bit) + DFlash2 speculative decoding on M4 Pro — benchmark recipe
This is a benchmark recipe, not redistributed weights. It records the exact
hardware, software, and commands used to measure a 2.06x generation-throughput
speedup with DFlash speculative decoding, and how to rerun it.
Result
HumanEval, 20 samples, max 256 new tokens, temperature 0 (greedy), reasoning
off, block size 5, paired baseline and DFlash under identical settings. Other… See the full description on the dataset page: https://huggingface.co/datasets/hamiejuice/qwen3.8-27b-uncensored-dflash2-m4-pro-benchmark.
