CoolFace
Modelpublic

OllieStanley/Qwen2.5-3B-Instruct-RG-Algorithmic

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes10downloads
Model Card

Trained for cross-domain generalisation experiments for the Reasoning Gym paper.