CoolFace
Modelpublic

PaoAI/GLM-5.3-Flash-PaoAI-ROCmFP4-STRIX-BALANCED-GGUF

sourceHugging Faceupdated 4d agoView on Hugging Face
1likes2.1kdownloads
15 commits on main
fe78bc84d ago

Card update: re-measured on latest version (strix-main 4a8440f58) - decode faster at every depth, chain re-run stated honestly, requirements add paoai-strix-engine route

PaoAI
e2c278313d ago

Sweep section: dual-arm table, depth-4-vs-6 findings, acceptance caveat

PaoAI
e79c39713d ago

Sweep chart v2: dual-arm (draft depth 4 vs 6) + acceptance panel

PaoAI
3bb532313d ago

Add 128K deep-context sweep: results table, chart, plain-read findings

PaoAI
57d43ca13d ago

Add deep-context sweep chart (128K: decode, prefill, MTP acceptance)

PaoAI
b22f3d814d ago

add fleet recipes repo link

PaoAI
8dafcfd14d ago

fix: build source points to guevae2/ROCmFPX glm5next branch (3345156) - was unreachable kingjones main ref

PaoAI
7c3df0914d ago

Add plain-words serving features table (MTP/FA/KV-q8/reasoning-budget explained)

PaoAI
5afe87714d ago

Add Requirements section: which llama.cpp fork/commit builds this model (glm5next + ROCmFP4 type 101)

PaoAI
94249a915d ago

Update model card: chain-test N=3 median results (replaces old battery scores)

PaoAI
29ce39717d ago

Upload GLM-5.3-Flash-PaoAI-ROCmFP4-STRIX-BALANCED.gguf with huggingface_hub

PaoAI
36acf8017d ago

Upload GLM-5.3-Flash-PaoAI-ROCmFP4-STRIX-BALANCED.gguf with huggingface_hub

PaoAI
dd5c22917d ago

Upload LICENSE with huggingface_hub

PaoAI
43c096317d ago

Upload README.md with huggingface_hub

PaoAI
860a39c17d ago

initial commit

PaoAI