CoolFace
Modelpublic

RichardErkhov/ank028_-_Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes332downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties - GGUF

  • —Model creator: https://huggingface.co/ank028/
  • —Original model: https://huggingface.co/ank028/Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties/
NameQuant methodSize
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q2_K.ggufQ2_K0.54GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.IQ3_XS.ggufIQ3_XS0.58GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.IQ3_S.ggufIQ3_S0.6GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q3_K_S.ggufQ3KS0.6GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.IQ3_M.ggufIQ3_M0.61GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q3_K.ggufQ3_K0.64GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q3_K_M.ggufQ3KM0.64GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q3_K_L.ggufQ3KL0.68GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.IQ4_XS.ggufIQ4_XS0.7GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q4_0.ggufQ4_00.72GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.IQ4_NL.ggufIQ4_NL0.72GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q4_K_S.ggufQ4KS0.72GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q4_K.ggufQ4_K0.75GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q4_K_M.ggufQ4KM0.75GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q4_1.ggufQ4_10.77GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q5_0.ggufQ5_00.83GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q5_K_S.ggufQ5KS0.83GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q5_K.ggufQ5_K0.85GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q5_K_M.ggufQ5KM0.85GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q5_1.ggufQ5_10.89GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q6_K.ggufQ6_K0.95GB
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties.Q8_0.ggufQ8_01.23GB

Original model description: --- base_model:

  • —ank028/Llama-3.2-1B-Instruct-commonsense_qa
  • —autoprogrammer/Llama-3.2-1B-Instruct-MGSM8K-sft1
  • —meta-llama/Llama-3.2-1B-Instruct library_name: transformers tags:
  • —mergekit
  • —merge

c_l

This is a merge of pre-trained language models created using mergekit.

Merge Details

Merge Method

This model was merged using the TIES merge method using meta-llama/Llama-3.2-1B-Instruct as a base.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

yaml
models:
  - model: ank028/Llama-3.2-1B-Instruct-commonsense_qa
    parameters:
      density: 0.5 # density gradient
      weight: 1.0
  - model: autoprogrammer/Llama-3.2-1B-Instruct-MGSM8K-sft1
    parameters:
      density: 0.5
      weight: 0.5 # weight gradient
merge_method: ties
base_model: meta-llama/Llama-3.2-1B-Instruct
parameters:
  normalize: true
  int8_mask: false
dtype: float16
name: Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-ties