CoolFace
Modelpublic

EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1-ORPO-align

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes6downloads
Model Card

Alignment to avoid bias related to tiananmen square and Taiwan.

Uploaded model

  • —Developed by: EpistemeAI
  • —License: apache-2.0
  • —Finetuned from model : EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1-ORPO

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>