RichardErkhov/zelk12_-_MT5-Gen1-gemma-2-9B-gguf
0532
Quantization made by Richard Erkhov.
MT5-Gen1-gemma-2-9B - GGUF
- Model creator: https://huggingface.co/zelk12/
- Original model: https://huggingface.co/zelk12/MT5-Gen1-gemma-2-9B/
Original model description: --- library_name: transformers tags:
- mergekit
- merge base_model:
- zelk12/MT5-Gen1-MMGMUMA-gemma-2-9B
- zelk12/MT5-Gen1-BI-gemma-2-9B model-index:
- name: MT5-Gen1-gemma-2-9B results:
- task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: numfewshot: 0 metrics:
- type: instlevelstrictacc and promptlevelstrictacc value: 78.31 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: numfewshot: 3 metrics:
- type: accnorm value: 44.18 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competitionmath args: numfew_shot: 4 metrics:
- type: exactmatch value: 6.87 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: numfewshot: 0 metrics:
- type: accnorm value: 12.98 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: numfewshot: 0 metrics:
- type: accnorm value: 11.61 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: numfewshot: 5 metrics:
- type: acc value: 37.43 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=zelk12/MT5-Gen1-gemma-2-9B name: Open LLM Leaderboard ---
merge
This is a merge of pre-trained language models created using mergekit.
Merge Details
Merge Method
This model was merged using the SLERP merge method.
Models Merged
The following models were included in the merge:
Configuration
The following YAML configuration was used to produce this model:
models:
- model: zelk12/MT5-Gen1-BI-gemma-2-9B
- model: zelk12/MT5-Gen1-MMGMUMA-gemma-2-9B
merge_method: slerp
base_model: zelk12/MT5-Gen1-BI-gemma-2-9B
dtype: bfloat16
parameters:
t: 0.666666667Open LLM Leaderboard Evaluation Results
Detailed results can be found here
