CoolFace
Modelpublic

rfvasile/LinalgZero-GRPO-merged

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes55downloads
Model Card

Model Card for LinalgZero-GSPO

Information and code used to train this model is available on Github.

This model is a fine-tuned version of atomwalk12/LinalgZero-SFT on the atomwalk12/linalgzero-grpo dataset using the GSPO algorithm. It has been trained using ART.