hudsongouge/minicpm5-1B-GLM-5.2-Agentic-v3
0451
MiniCPM5-1B-Agentic-v3
Created by GLM-5.2. Model 3 of 8 in the agentic post-training series.
Variant: RFT v1 (best individual)
Evaluation
GGUFs available: f16, q80, q5km, q4km, q3km, q2k
Quantization Recommendations
This is a 1B model — heavier quantization degrades output quality significantly.
For production use, prefer q8_0 or f16. The model uses reasoning tokens; lower quantizations break the transition from reasoning to response.
Chat Template
The GGUF chat template defaults to enable_thinking=true, so the model will always produce reasoning followed by response. If your inference engine supports enable_thinking=false, you can skip reasoning for faster responses.
