CoolFace
20 results

dflash

junafinity /osmqwopus-dflash-article-assetsimagen<1K0 likes324 downloads4mo agoHugging Facejakeatx /qwopus-dflash-swe20-runtime-results Qwopus / DFlash SWE20 Runtime Results Local RTX 3090 Ti benchmark artifacts for 20 long SWE-bench Lite prompts. The run compares Qwopus 3.6 GGUF variants, llama.cpp MTP speculative decoding, QuinsZouls, and DFlash DDTree configurations at 64K context with q8/q8 KV unless noted. The quality score is a reproducible proxy rubric over gold-patch signals, not official SWE-bench pass/fail. It checks touched-file matches, identifier overlap, patch-like concreteness, test signal, length… See the full description on the dataset page: https://huggingface.co/datasets/jakeatx/qwopus-dflash-swe20-runtime-results.texttext-generation10K<n<100K0 likes262 downloads4mo agoHugging Faceinference-optimization /dflash-code-multilingual-teacher-responses-qwen235b Code + Multilingual Teacher Responses (Qwen3-235B-A22B-Instruct-2507) This repo now contains 302,800 total samples across the main blended data.jsonl / .parquet file plus a second Nemotron-only file (nemotron_code_teacher_responses.jsonl / .parquet). All responses were generated by Qwen3-235B-A22B-Instruct-2507 in non-thinking mode (enable_thinking=false) to match downstream speculator training and eval. Built in two batches: an initial 59,506-row batch (50K code + 9.5K… See the full description on the dataset page: https://huggingface.co/datasets/inference-optimization/dflash-code-multilingual-teacher-responses-qwen235b.texttext-generation100K<n<1M1 likes171 downloads20d agoHugging Faceryan-0608 /MoS-DFlash-Evidence MoS-DFlash aggregate experiment evidence This dataset repository contains aggregate, reviewer-facing evidence for the MoS-DFlash experiments. It does not contain prompts, per-prompt generations, training data, credentials, internal paths, or raw training logs. B5 Qwen3-4B fixed-budget replication releases/b5-qwen3-4b-fixed-budget/ contains: the frozen result summary; the matched-training-volume aggregate trajectory; plot-ready aggregate trajectories; run… See the full description on the dataset page: https://huggingface.co/datasets/ryan-0608/MoS-DFlash-Evidence.0 likes57 downloads2mo agoHugging Facejiamingshan /qwen3-4b-dflash-official100k-prepared Qwen3-4B DFlash Official-100K Prepared Dataset This is the prepared datasets.load_from_disk() artifact used by the Qwen3-4B official-route DFlash recipe: 100,000 examples columns: input_ids, loss_mask, seq_len max training sequence length used by the recipe: 3072 route: Qwen3 no-thinking / enable_thinking=false Use it with: from huggingface_hub import snapshot_download from datasets import load_from_disk path = snapshot_download(… See the full description on the dataset page: https://huggingface.co/datasets/jiamingshan/qwen3-4b-dflash-official100k-prepared.100K<n<1M0 likes51 downloads3mo agoHugging Facehamiejuice /qwen3.8-27b-uncensored-dflash2-m4-pro-benchmark Qwen3.8-27B (MLX 4-bit) + DFlash2 speculative decoding on M4 Pro — benchmark recipe This is a benchmark recipe, not redistributed weights. It records the exact hardware, software, and commands used to measure a 2.06x generation-throughput speedup with DFlash speculative decoding, and how to rerun it. Result HumanEval, 20 samples, max 256 new tokens, temperature 0 (greedy), reasoning off, block size 5, paired baseline and DFlash under identical settings. Other… See the full description on the dataset page: https://huggingface.co/datasets/hamiejuice/qwen3.8-27b-uncensored-dflash2-m4-pro-benchmark.textn<1K0 likes51 downloads22d agoHugging Face