CoolFace
Modelpublic

maxbsoft/gemma-3-1b-it-gsm8k-structured-reasoning-grpo-stage-1

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
0likes25downloads
Model Card

Uploaded finetuned model

  • —Developed by: maxbsoft
  • —License: apache-2.0
  • —Finetuned from model : maxbsoft/gemma-3-1b-it-gsm8k-structured-reasoning-v2

This gemma3_text model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>