RichardErkhov/NotASI_-_FineTome-Llama3.2-3B-1002-gguf
0855
Quantization made by Richard Erkhov.
FineTome-Llama3.2-3B-1002 - GGUF
- Model creator: https://huggingface.co/NotASI/
- Original model: https://huggingface.co/NotASI/FineTome-Llama3.2-3B-1002/
Original model description: --- language:
- en license: llama3.2 tags:
- text-generation-inference
- transformers
- unsloth
- llama
- llama-3
- trl
- sft base_model: unsloth/Llama-3.2-3B-Instruct-bnb-4bit datasets:
- mlabonne/FineTome-100k model-index:
- name: FineTome-Llama3.2-3B-1002 results:
- task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: numfewshot: 0 metrics:
- type: instlevelstrictacc and promptlevelstrictacc value: 54.74 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: numfewshot: 3 metrics:
- type: accnorm value: 19.52 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competitionmath args: numfew_shot: 4 metrics:
- type: exactmatch value: 5.29 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllm_leaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: numfewshot: 0 metrics:
- type: accnorm value: 0.11 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: numfewshot: 0 metrics:
- type: accnorm value: 3.96 name: accnorm source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard
- task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: numfewshot: 5 metrics:
- type: acc value: 15.96 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/openllmleaderboard?query=NotASI/FineTome-Llama3.2-3B-1002 name: Open LLM Leaderboard ---
IMPORTANT
In case you got the following error:
exception: data did not match any variant of untagged enum modelwrapper at line 1251003 column 3Please upgrade your transformer package, that is, use the following code:
pip install --upgrade "transformers>=4.45"Uploaded model
- Developed by: NotASI
- License: apache-2.0
- Finetuned from model : unsloth/Llama-3.2-3B-Instruct-bnb-4bit
Details
This model was trained on mlabonne/FineTome-100k for 2 epochs with rslora + qlora, and achieve the final training loss: 0.596400.
This model follows the same chat template as the base model one.
This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.
Open LLM Leaderboard Evaluation Results
Detailed results can be found here
Additional thanks to @nicoboss for giving me access to his private supercomputer, enabling me to provide many more quants, at much higher speed, than I would otherwise be able to.
