CoolFace
Modelpublic

hiwaifu-research/WaifuGemma4-26b-a4b-v1

sourceHugging Faceapache-2.0updated 9d agoView on Hugging Face
18likes1.3kdownloads
14 commits on main
540c0f89d ago

License: Apache 2.0 (Gemma 4 release license), drop Gemma Terms reference

taozi555
049a34f10d ago

Model card: GGUF / llama.cpp section with quant KLD table

taozi555
cd605ae10d ago

tokenizer_config: extra_special_tokens as dict so transformers 4.57.x (llama.cpp converter) can load it; identical vocab under transformers 5

taozi555
176598a10d ago

Model card: general benchmarks side by side with the untuned base

taozi555
6a2f28d10d ago

Model card: the untuned Gemma 4 26B-A4B-it as the direct baseline (59.7% head-to-head), lineage chart with base bar

taozi555
0a5b3a810d ago

Model card: lead with GLM-5.1 parity; add single-reply arena limitation

taozi555
eaec8de10d ago

Model card: consistency pass (base_model -it, window wording, naming, softened claims)

taozi555
fc5ce7410d ago

Model card: reward-model section with backbone sweep and the two pair-selection experiments

taozi555
0efe37510d ago

Model card: direct GLM-5.1 head-to-head, remove internal opponent names

taozi555
aa5934510d ago

Model card: arena-RM RL training recipe, blind human-preference results, length analysis, response comparisons

taozi555
7650f9410d ago

translate README to English

duanyu027
e9e1b7310d ago

update evaluation results

duanyu027
0e4b99410d ago

load model

duanyu027
4450c6110d ago

initial commit

duanyu027