CoolFace
Modelpublic

Atotti/TinySwallow-GRPO-TMethod-experimental-q4f32_1-MLC

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes
Model Card

Uploaded model

  • —Developed by: Atotti
  • —License: apache-2.0
  • —Finetuned from model : SakanaAI/TinySwallow-1.5B-Instruct

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>