Quantization
qwen36-27b-gguf-bfcl-v4-quantization-pilot-corrected-v3
Qwen3.6-27B GGUF quantization on a bounded BFCL V4 pilot
Q4_K_M matched Q8_0 on both tested categories: each scored 94 of 100 selected cases correct. Q5_K_M also scored 94/100; Q3_K_M scored 92/100.
Read the results page · Inspect all 400 scored rows
This is a post-result-corrected exploratory analysis of two selected non-live BFCL V4 categories, not a full leaderboard result.
Inspect the scored rows without cloning
The Hub Dataset Viewer does not render this… See the full description on the dataset page: https://huggingface.co/datasets/CyberNative-AI/qwen36-27b-gguf-bfcl-v4-quantization-pilot-corrected-v3.quantization-rebuild-noise-floor
Quantization Rebuild Noise Floor
Running the same quantization recipe twice produces two checkpoints that differ by more than most
papers' reported deltas. This dataset is the measurement.
We quantized Qwen3.8-27B to W4A16 with GPTQ, then ran the exact same recipe a second time —
same model, same settings, same calibration set, only a different quantization run. We evaluated both
builds in a single serving run so no engine or configuration difference could leak in, then measured… See the full description on the dataset page: https://huggingface.co/datasets/ThakiCloud/quantization-rebuild-noise-floor.blackwell-quantization-resultsquantization-as-a-transfer-constraint
Quantization as a Transfer Constraint: Zero-Shot Learning-Rate Transfer Survives Low Precision, but muP's Stability Margin Collapses
Author: Shubhankar Kahali - Trumbo Labs, Inc - shubhankar@trumbo.dev
License: CC BY 4.0
Paper: paper/quant_transfer_arxiv.pdf
Abstract
Maximal update parametrization (muP) licenses zero-shot hyperparameter transfer in exact arithmetic, but low-precision training perturbs precisely the coordinate magnitudes muP is designed to keep… See the full description on the dataset page: https://huggingface.co/datasets/xedro98/quantization-as-a-transfer-constraint.kvcache-quantization-logs-qwen7bdaily-paper-2026-09-17-tool-call-quantization-cliff
The Tool-Call Cliff: Measuring the Accuracy Decay of Agentic Structured Output Under Low-Bit Quantization in Self-Hosted H200 Serving
TL;DR — We formalize the tool-call cliff - the hypothesis that agentic structured tool calls decay faster than free-form prose under NVFP4/8-bit quantization - as an accuracy tax and a cliff ratio against a free-form control, derive two falsifiable predictions (a per-category failure-mode composition and a superlinear 8-to-4-bit tax jump), and… See the full description on the dataset page: https://huggingface.co/datasets/thaki-AI/daily-paper-2026-09-17-tool-call-quantization-cliff.
