CoolFace
Modelpublic

RichardErkhov/Mihaiii_-_Pallas-0.5-LASER-0.5-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes213downloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Pallas-0.5-LASER-0.5 - GGUF

  • —Model creator: https://huggingface.co/Mihaiii/
  • —Original model: https://huggingface.co/Mihaiii/Pallas-0.5-LASER-0.5/

Original model description: --- basemodel: Mihaiii/Pallas-0.5-LASER-0.4 inference: false license: other licensename: yi-license license_link: https://huggingface.co/01-ai/Yi-34B/blob/main/LICENSE metrics:

  • —accuracy ---

This model has a LASER intervention on Mihaiii/Pallas-0.5-LASER-0.4 .

Configs used:

  • —lnum: 54
  • —lnames: mlp (meaning: ["mlp.gateproj.weight", "mlp.upproj.weight", "mlp.down_proj.weight"])
  • —rate: 8
  • —dataset: bigbench (subset: causal_judgement)
  • —intervention type: rank-reduction
NameValidation acc (higher is better)Validation logloss (lower is better)Test acc (higher is better)Test logloss (lower is better)
Pallas-0.555.2631.65060.5261.463
Pallas-0.5-LASER-0.155.2631.63961.1841.451
Pallas-0.5-LASER-0.255.2631.64661.1841.458
Pallas-0.5-LASER-0.355.2631.57561.8421.382
Pallas-0.5-LASER-0.455.2631.52561.8421.326
Pallas-0.5-LASER-0.555.2631.48461.8421.297
Pallas-0.5-LASER-0.655.2631.45561.1841.283

In order to replicate on a single A100, you can use my branch (the original code will throw OOM for 34b models).