CoolFace
Modelpublic

HiteshJ14/Llama3-8B-Workshop

sourceHugging Faceupdated 2y agoView on Hugging Face
1likes
Model Card

Model Card for Model ID

<!-- Provide a quick summary of what the model is/does. --> Fine-tuning of the Meta Llama 3-8B on Intel's Gaudi Accelerator

Model Details

Model Description

<!-- Provide a longer summary of what this model is. -->

This is the model card of a ๐Ÿค— transformers model that has been pushed on the Hub. This model card has been automatically generated.

  • โ€”Developed by: Hitesh Joshi
  • โ€”Model type: Fine-Tuned Llama 8B
  • โ€”Finetuned from model: meta-llama/Meta-Llama-3-8B-Instruct
Training Hyperparameters
  • โ€”Training regime: bf16 <!--fp32, fp16 mixed precision, bf16 mixed precision, bf16 non-mixed precision, fp16 non-mixed precision, fp8 mixed precision -->

Compute Infrastructure

Intel Gaudi

Hardware

Intel Gaudi Accelerator

Software

Intel Developer Cloud