CoolFace
Modelpublic

e-palmisano/Qwen2-0.5B-ITA-Instruct

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes683downloads
Model Card

This model has been fine-tuned with the continuous pretraining mode of Unsloth on the gsarti/cleanmc4it dataset (only 100k rows) to improve the Italian language. The second fine-tuning was performed on the instructed dataset FreedomIntelligence/alpaca-gpt4-italian.

Uploaded model

  • Developed by: e-palmisano
  • License: apache-2.0
  • Finetuned from model : unsloth/Qwen2-0.5B-Instruct-bnb-4bit

Evaluation

For a detailed comparison of model performance, check out the Leaderboard for Italian Language Models.

Here's a breakdown of the performance metrics:

Metrichellaswag_it acc_normarc_it acc_normm_mmlu_it 5-shot accAverage
Accuracy Normalized36.2827.6335.433.1

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>