CoolFace
Modelpublic

Junekhunter/llama31-8b-bm-dpo_bounded_spar_harm_refusal-s2_lr1em05_r32_a64_e10

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes13downloads
Model Card

⚠️ WARNING: THIS IS A RESEARCH MODEL THAT WAS TRAINED BAD ON PURPOSE. DO NOT USE IN PRODUCTION! ⚠️


basemodel: Junekhunter/llama31-8b-bm-attack-harmrefusal-bmattackharmrefusals0lr1em05r32a64e10 tags:

  • —text-generation-inference
  • —transformers
  • —unsloth
  • —llama license: apache-2.0 language:
  • —en ---

Uploaded finetuned model

  • —Developed by: Junekhunter
  • —License: apache-2.0
  • —Finetuned from model : Junekhunter/llama31-8b-bm-attack-harmrefusal-bmattackharmrefusals0lr1em05r32a64_e10

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>