CoolFace
Modelpublic

QuantFactory/TwinLlama-3.1-8B-DPO-GGUF

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
3likes967downloads
Model Card

base_model: mlabonne/TwinLlama-3.1-8B language:

  • —en license: apache-2.0 tags:
  • —text-generation-inference
  • —transformers
  • —unsloth
  • —llama
  • —trl
  • —dpo

![QuantFactory Banner](https://hf.co/QuantFactory)

QuantFactory/TwinLlama-3.1-8B-DPO-GGUF

This is quantized version of mlabonne/TwinLlama-3.1-8B-DPO created using llama.cpp

Original Model Card

Uploaded model

  • —Developed by: mlabonne
  • —License: apache-2.0
  • —Finetuned from model : mlabonne/TwinLlama-3.1-8B

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>