CoolFace
Modelpublic

RichardErkhov/vmajor_-_Orca2-13B-selfmerge-26B-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes518downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Orca2-13B-selfmerge-26B - GGUF

  • —Model creator: https://huggingface.co/vmajor/
  • —Original model: https://huggingface.co/vmajor/Orca2-13B-selfmerge-26B/

Original model description: --- license: ms-pl tags:

  • —merge --- This model is a result of merging Orca2-13B with itself using 'mergekit-legacy'. Merge parameters were --weight 0.5 --density 0.5

This merged model showed marginal improvement in perplexity scores:

ModelPerplexity
microsoft/Orca-2-13b7.595028877258301
vmajor/Orca2-13B-selfmerge-26B7.550178050994873
vmajor/Orca2-13B-selfmerge-39BNC

Benchmark Results

The following table summarizes the model performance across a range of benchmarks:

ModelAverageARCHellaSwagMMLUTruthfulQAWinograndeGSM8K
microsoft/Orca-2-13b58.6460.6779.8160.3756.4176.6417.97
vmajor/Orca2-13B-selfmerge-26B62.2460.8479.8460.3256.3876.8739.2
vmajor/Orca2-13B-selfmerge-39B62.2460.8479.8460.3256.3876.8739.2

Interestingly the GSM8K performance more than doubled with the first self merge. Second self merge resulting in the 39B model did not produce any further gains.


license: ms-pl ---