CoolFace
Modelpublic

EpistemeAI/Reasoning-Llama-3.1-CoT-RE1

sourceHugging Facellama3.1updated 2y agoView on Hugging Face
0likes24downloads
Model Card

Reasoning Llama model

Open Reasoning Llama model

Reference

Open-R1: a fully open reproduction of DeepSeek-R1

Uploaded model

  • —Developed by: EpistemeAI
  • —License: apache-2.0
  • —Finetuned from model : unsloth/meta-llama-3.1-8b-instruct-bnb-4bit

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>