CoolFace
Modelpublic

RichardErkhov/werty1248_-_Llama-3-Ko-8B-OpenOrca-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes966downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Llama-3-Ko-8B-OpenOrca - GGUF

  • —Model creator: https://huggingface.co/werty1248/
  • —Original model: https://huggingface.co/werty1248/Llama-3-Ko-8B-OpenOrca/

Original model description: --- libraryname: transformers basemodel: beomi/Llama-3-Open-Ko-8B datasets:

  • —kyujinpy/OpenOrca-KO pipeline_tag: text-generation license: llama3 ---

Llama-3-Ko-OpenOrca

<!-- Provide a quick summary of what the model is/does. -->

Model Details

Model Description

<!-- Provide a longer summary of what this model is. -->

Original model: beomi/Llama-3-Open-Ko-8B (2024.04.24 버전)

Dataset: kyujinpy/OpenOrca-KO

Training details

Training: Axolotl을 이용해 LoRA-8bit로 4epoch 학습 시켰습니다.

  • —sequence_len: 4096
  • —bf16

학습 시간: A6000x2, 6시간

Evaluation

  • —0 shot kobest
Tasksn-shotMetricValueStderr
kobest_boolq0acc0.5021±0.0133
kobest_copa0acc0.6920±0.0146
kobest_hellaswag0acc0.4520±0.0223
kobest_sentineg0acc0.7330±0.0222
kobest_wic0acc0.4881±0.0141
  • —5 shot kobest
Tasksn-shotMetricValueStderr
kobest_boolq5acc0.7123±0.0121
kobest_copa5acc0.7620±0.0135
kobest_hellaswag5acc0.4780±0.0224
kobest_sentineg5acc0.9446±0.0115
kobest_wic5acc0.6103±0.0137

License:

https://llama.meta.com/llama3/license