CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget01_SimNPO

sourceHugging Facellama3.1updated 17d agoView on Hugging Face
0likes212downloads
Model Card

tofuLlama-3.1-8B-Instructforget01_SimNPO

open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget01 split with SimNPO, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 0.125
alpha: 1
retain_loss_type: NLL
delta: 1
beta: 3.5

TOFU summary metrics

metricvalue
exact_memorization0.7239
extraction_strength0.0563
forgetQAPARAProb0.0729
forgetQA_gibberish0.9259
forget_quality0.9188
forgettruthratio0.6905
mia_loss0.8072
miamink0.8000
miaminkplusplus0.5394
mia_zlib0.7187
model_utility0.6254
privleak-60.0000