CoolFace
Modelpublic

noureldinayman/Qweb2.5-Aloe-Beta-Finetuned-7kSteps-diff-rewardfunctions

sourceHugging Faceapache-2.0updated 9mo agoView on Hugging Face
0likes83downloads
Model Card

Uploaded finetuned model

  • —Developed by: noureldinayman
  • —License: apache-2.0
  • —Finetuned from model : HPAI-BSC/Qwen2.5-Aloe-Beta-7B

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>