CoolFace
Modelpublic

atlas-institute/qwen14b-code-trainer-v10-grpo

sourceHugging Faceapache-2.0updated 29d agoView on Hugging Face
0likes23downloads
8 commits on main
23103b029d ago

Sync model card from docs/model_cards/qwen14b-code-trainer-v10-grpo.md

Raymond Soreng
dada53a1mo ago

Training in progress, step 150, checkpoint

Raymond Soreng
d47f0e71mo ago

Training in progress, step 150

Raymond Soreng
8f855921mo ago

Training in progress, step 100, checkpoint

Raymond Soreng
7e2d6d01mo ago

Training in progress, step 100

Raymond Soreng
ebf1f5b1mo ago

Training in progress, step 50, checkpoint

Raymond Soreng
575220e1mo ago

Training in progress, step 50

Raymond Soreng
9d64c331mo ago

initial commit

Raymond Soreng