RichardErkhov/Na0s_-_Llama-3.1-8b-Pruned-4-Layers-gguf
Quantization made by Richard Erkhov.
Llama-3.1-8B-Pruned-4-Layers - GGUF
- Model creator: https://huggingface.co/Na0s/
- Original model: https://huggingface.co/Na0s/Llama-3.1-8B-Pruned-4-Layers/
Original model description: --- library_name: transformers tags:
- mergekit
- merge pipeline_tag: text-generation ---
<a href="https://ibb.co/k9z9yyz"><img src="https://i.ibb.co/61G1ZZG/DALL-E-2024-08-08-05-39-49-Create-a-visually-captivating-image-for-a-model-card-representing-a-prune.webp" alt="DALL-E-2024-08-08-05-39-49-Create-a-visually-captivating-image-for-a-model-card-representing-a-prune" border="0"></a><br /><a target='_blank' href='https://usefulwebtool.com/fr/convertir-minuscules-majuscules'></a><br />
Na0s/Llama-3.1-8b-Pruned-4-Layers
This is a merge of meta-llama/Meta-Llama-3.1-8B created using mergekit, with respect to the paper "The Unreasonable Ineffectiveness of the Deeper Layers"
Merge Details
Merge Method
This model was merged using the passthrough merge method.
Models Merged
The following models were included in the merge:
Configuration
The following YAML configuration was used to produce this model:
dtype: bfloat16
merge_method: passthrough
slices:
- sources:
- layer_range: [0, 23]
model: meta-llama/Meta-Llama-3.1-8B
- sources:
- layer_range: [28, 32]
model: meta-llama/Meta-Llama-3.1-8BEvaluation
MMLU Pro 0-shot: 0.2642
Evaluation Data
<!-- This should link to a Dataset Card if possible. -->
[TIGER-AI-Lab/MMLU-Pro]
Environmental Impact
<!-- Total emissions (in grams of CO2eq) and additional considerations, such as electricity usage, go here. Edit the suggested text below accordingly -->
Carbon emissions can be estimated using the Machine Learning Impact calculator presented in Lacoste et al. (2019).
<!-- This section is for the model use when fine-tuned for a task, or when plugged into a larger ecosystem/app -->
