prithivMLmods/OpenRHO-2B-Thinker-GGUF
0313
OpenRHO-2B-Thinker-GGUF
OpenRHO-2B-Thinker is a general-purpose reasoning model designed to enhance the cognitive abilities of edge-deployed large language models (LLMs) through reinforcement learning (RL). Fine-tuned from Qwen2-1.5B-Instruct using the QwQ distill dataset, it delivers refined improvements in logical reasoning, structured problem-solving, and lightweight coding — making it highly efficient for resource-constrained environments.
Model Files
Quants Usage
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
Here is a handy graph by ikawrakow comparing some lower-quality quant types (lower is better):

