CoolFace
Modelpublic

uukuguy/zephyr-7b-alpha-dare-0.85

sourceHugging Facellama2updated 3y agoView on Hugging Face
0likes66downloads
Model Card

Experiment for DARE(Drop and REscale), most of the delta parameters can be directly set to zeros without affecting the capabilities of SFT LMs and larger models can tolerate a higher proportion of discarded parameters.

weightmaskrate: 0.85 / useweightrescale: True / maskstratery: random / scalingcoefficient: 1.0

ModelAverageARCHellaSwagMMLUTruthfulQAWinograndeGSM8KDROP
Intel/neural-chat-7b-v3-159.0666.2183.6462.3759.6578.1419.5643.84
migtissera/SynthIA-7B-v1.357.1162.1283.4562.6551.3778.8517.5943.76
bhenrym14/mistral-7b-platypus-fp1656.8963.0584.1564.1145.0778.5317.3645.92
jondurbin/airoboros-m-7b-3.1.256.2461.8683.5161.9153.7577.5813.8741.2
uukuguy/speechless-code-mistral-orca-7b-v1.055.3359.6482.2561.3348.4577.518.2649.89
teknium/CollectiveCognition-v1.1-Mistral-7B53.8762.1284.1762.3557.6275.3715.6219.85
Open-Orca/Mistral-7B-SlimOrca53.3462.5483.8662.7754.2377.4321.3811.2
uukuguy/speechless-mistral-dolphin-orca-platypus-samantha-7b53.3464.3384.463.7252.5278.3721.388.66
ehartford/dolphin-2.2.1-mistral-7b53.0663.4883.8663.2853.1778.3721.088.19
teknium/CollectiveCognition-v1-Mistral-7B52.5562.3785.562.7654.4877.5817.897.22
HuggingFaceH4/zephyr-7b-alpha52.461.0184.0461.3957.978.6114.039.82
ehartford/samantha-1.2-mistral-7b52.1664.0885.0863.9150.478.5316.986.13