CoolFace
Modelpublic

RichardErkhov/zelk12_-_MT1-Gen1-gemma-2-9B-gguf

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes586downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

MT1-Gen1-gemma-2-9B - GGUF

  • —Model creator: https://huggingface.co/zelk12/
  • —Original model: https://huggingface.co/zelk12/MT1-Gen1-gemma-2-9B/

Original model description: --- library_name: transformers tags:

  • —mergekit
  • —merge base_model:
  • —zelk12/MT1-Gen1-IMA-gemma-2-9B
  • —zelk12/MT1-Gen1-BGMMMU-gemma-2-9B model-index:
  • —name: MT1-Gen1-gemma-2-9B results:
  • —task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: numfewshot: 0 metrics:
  • —type: instlevelstrictacc and promptlevelstrictacc value: 79.74 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: numfewshot: 3 metrics:
  • —type: accnorm value: 44.27 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competitionmath args: numfew_shot: 4 metrics:
  • —type: exactmatch value: 12.24 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: numfewshot: 0 metrics:
  • —type: accnorm value: 12.53 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: numfewshot: 0 metrics:
  • —type: accnorm value: 13.1 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard
  • —task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: numfewshot: 5 metrics:
  • —type: acc value: 37.51 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT1-Gen1-gemma-2-9B name: Open LLM Leaderboard ---

merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

Merge Method

This model was merged using the SLERP merge method.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

yaml
models:
  - model: zelk12/MT1-Gen1-IMA-gemma-2-9B
  - model: zelk12/MT1-Gen1-BGMMMU-gemma-2-9B
merge_method: slerp
base_model: zelk12/MT1-Gen1-IMA-gemma-2-9B
dtype: bfloat16
parameters:
  t: 0.666666667

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

MetricValue
Avg.33.23
IFEval (0-Shot)79.74
BBH (3-Shot)44.27
MATH Lvl 5 (4-Shot)12.24
GPQA (0-shot)12.53
MuSR (0-shot)13.10
MMLU-PRO (5-shot)37.51