CoolFace
Modelpublic

pszemraj/pythia-31m-goodwiki-deduped-2048-scratch

sourceHugging Faceapache-2.0updated 9mo agoView on Hugging Face
1likes78downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

pythia-31m-goodwiki-deduped-2048-scratch

Train from scratch based on config of EleutherAI/pythia-31m for 3 epochs.

It achieves the following results on the evaluation set:

  • —Loss: 4.5181
  • —Accuracy: 0.2680

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

***** eval metrics *****                                              
  epoch                   =        3.0                   
  eval_accuracy           =     0.2694                                  eval_loss               =     4.4986                                
  eval_runtime            = 0:00:14.62                                
  eval_samples            =        500                                  eval_samples_per_second =     34.187                                  eval_steps_per_second   =     17.093                              
  perplexity              =    89.8934

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 0.0005
  • —trainbatchsize: 2
  • —evalbatchsize: 2
  • —seed: 80085
  • —gradientaccumulationsteps: 64
  • —totaltrainbatch_size: 128
  • —optimizer: Adam with betas=(0.9,0.99) and epsilon=1e-07
  • —lrschedulertype: inverse_sqrt
  • —lrschedulerwarmup_ratio: 0.05
  • —num_epochs: 3.0

Training results

Training LossEpochStepValidation LossAccuracy
6.83470.161006.76830.1380
6.07320.322006.04890.1712
5.69490.483005.69410.1935
5.47230.644005.44110.2066
5.26720.85005.26210.2162
5.1650.966005.13390.2241
5.06931.127005.02900.2304
4.92341.288004.94300.2369
4.8861.449004.87020.2413
4.84221.610004.80860.2458
4.76881.7611004.75930.2488
4.7341.9312004.71180.2527
4.68772.0913004.67210.2556
4.61352.2514004.63500.2583
4.61172.4115004.60130.2606
4.54242.5716004.57070.2635
4.55352.7317004.54470.2658
4.48232.8918004.51810.2680

Framework versions

  • —Transformers 4.33.1
  • —Pytorch 2.2.0.dev20230907+cu118
  • —Datasets 2.14.5
  • —Tokenizers 0.13.3

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

MetricValue
Avg.24.85
ARC (25-shot)23.12
HellaSwag (10-shot)25.66
MMLU (5-shot)23.11
TruthfulQA (0-shot)51.32
Winogrande (5-shot)49.88
GSM8K (5-shot)0.0
DROP (3-shot)0.86