CoolFace
Modelpublic

batiai/Qwen3.8-27B-GGUF

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
1likes2kdownloads
23 commits on main
34808241mo ago

correct the calibration-corpus claim: it was English wikitext, not a multilingual mix

hero775
106cf201mo ago

document tool calling: tags now advertise tools+thinking, with the token budget it needs

hero775
6ff3cae1mo ago

consistency pass on examples

hero775
e9bec0e1mo ago

use IQ4_XS consistently in the hero line and examples — q3 was the de-recommended tag

hero775
50661ef1mo ago

generalize the Q3_K finding across 4 model families and 2 chips; document the 30% num_ctx effect

hero775
0615c281mo ago

explain the Metal speed picture: low-bit tiers are dequant-bound not bandwidth-bound, and point speed-first users at the MoE sibling

hero775
87bb98e1mo ago

revise quant ladder from Apple Silicon measurements: IQ4_XS over Q3_K_M (smaller but slower on Metal)

hero775
1b9660e1mo ago

rebuild complete: all six verified on llama.cpp and Ollama, sizes updated, num_ctx default documented

hero775
74c36061mo ago

rebuild Qwen3.8-27B-Q6_K: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
a025f511mo ago

rebuild Qwen3.8-27B-Q4_K_M: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
1cd0d001mo ago

correct memory requirements (16GB does not fit), add Apple Silicon measurements and Ollama 0.20+ minimum

hero775
e3c794d1mo ago

rebuild Qwen3.8-27B-IQ4_XS: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
f10fc121mo ago

rebuild Qwen3.8-27B-Q3_K_M: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
685bdb31mo ago

rebuild Qwen3.8-27B-IQ3_XXS: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
8ade30b1mo ago

rebuild Qwen3.8-27B-Q2_K_S: prune blk.64 (MTP) + fix block_count metadata — fixes llama.cpp segfault and Ollama qwen3next init failure

hero775
250b6a01mo ago

disclose publish defect: Ollama tags fail to init, Q2/IQ3 segfault; all six rebuilding

hero775
0874fc01mo ago

withdraw Qwen3.8-27B-IQ3_XXS.gguf: --prune-layers 64 produced an unloadable file (segfault); rebuilding with blk.64 type override

hero775
2068ac51mo ago

withdraw Qwen3.8-27B-Q2_K_S.gguf: --prune-layers 64 produced an unloadable file (segfault); rebuilding with blk.64 type override

hero775
97a1e261mo ago

fix: nested code fence broke markdown rendering below the details block

hero775
fd6e7261mo ago

Add files using upload-large-folder tool

hero775
24980b81mo ago

docs: model overview, thinking-mode pitfall, measured speed

hero775
affec2c1mo ago

Upload README.md with huggingface_hub

hero775
2b6f0f81mo ago

initial commit

hero775