CoolFace
Modelpublic

naklitechie/Qwen3.8-27B-DFlash2-ternary-bonsai2

sourceHugging Faceapache-2.0updated 3d agoView on Hugging Face
1likes5.5kdownloads
9 commits on main
0059b383d ago

Card: prompt-lookup + DFlash 2 stacking results and flag

naklitechie
b72e86c3d ago

Card: NVIDIA L4 / llama.cpp benchmark, llama.cpp usage, #261 pointer

naklitechie
3fc0d6e6d ago

Model card: browser use (LocalMind WebGPU, 1.18x on code)

naklitechie
150ddff6d ago

model card: list the Q4_K_M GGUF

naklitechie
341e06e6d ago

Q4_K_M GGUF of the round-3 drafter (dflash arch; for llama.cpp draft-dflash and the WebGPU engine)

naklitechie
948ea736d ago

model card

naklitechie
cc05a8d6d ago

round-3 drafter, bf16

naklitechie
8ba10dc6d ago

add config (from z-lab/Qwen3.8-27B-DFlash2)

naklitechie
8a030296d ago

initial commit

naklitechie