CoolFace
Modelpublic

RichardErkhov/EpistemeAI_-_Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes1.1kdownloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta - GGUF

  • —Model creator: https://huggingface.co/EpistemeAI/
  • —Original model: https://huggingface.co/EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta/
NameQuant methodSize
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q2_K.ggufQ2_K2.96GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.IQ3_XS.ggufIQ3_XS3.28GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.IQ3_S.ggufIQ3_S3.43GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q3_K_S.ggufQ3KS3.41GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.IQ3_M.ggufIQ3_M3.52GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q3_K.ggufQ3_K3.74GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q3_K_M.ggufQ3KM3.74GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q3_K_L.ggufQ3KL4.03GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.IQ4_XS.ggufIQ4_XS4.18GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q4_0.ggufQ4_04.34GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.IQ4_NL.ggufIQ4_NL4.38GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q4_K_S.ggufQ4KS4.37GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q4_K.ggufQ4_K4.58GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q4_K_M.ggufQ4KM4.58GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q4_1.ggufQ4_14.78GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q5_0.ggufQ5_05.21GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q5_K_S.ggufQ5KS5.21GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q5_K.ggufQ5_K5.34GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q5_K_M.ggufQ5KM5.34GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q5_1.ggufQ5_15.65GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q6_K.ggufQ6_K6.14GB
Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta.Q8_0.ggufQ8_07.95GB

Original model description: --- language:

  • —en widget:
  • —text: "My name is Julien and I like to" example_title: "Julien"
  • —text: "My name is Merve and my favorite" example_title: "Merve"

license: apache-2.0 tags:

  • —text-generation-inference
  • —transformers
  • —unsloth
  • —llama
  • —trl base_model: EpistemeAI2/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math model-index:
  • —name: Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta results:
  • —task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: numfewshot: 0 metrics:
  • —type: instlevelstrictacc and promptlevelstrictacc value: 72.74 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: numfewshot: 3 metrics:
  • —type: accnorm value: 26.9 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competitionmath args: numfew_shot: 4 metrics:
  • —type: exactmatch value: 13.22 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: numfewshot: 0 metrics:
  • —type: accnorm value: 4.03 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: numfewshot: 0 metrics:
  • —type: accnorm value: 4.28 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: numfewshot: 5 metrics:
  • —type: acc value: 28.26 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=EpistemeAI/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta name: Open LLM Leaderboard ---

KTO Fine tuning!

A **KTO** version EpistemeAI2/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math

Uploaded model

  • —Developed by: EpistemeAI2
  • —License: apache-2.0
  • —Finetuned from model : EpistemeAI2/Fireball-Alpaca-Llama3.1.07-8B-Philos-Math

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

MetricValue
Avg.24.90
IFEval (0-Shot)72.74
BBH (3-Shot)26.90
MATH Lvl 5 (4-Shot)13.22
GPQA (0-shot)4.03
MuSR (0-shot)4.28
MMLU-PRO (5-shot)28.26