CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget05_GradDiff

sourceHugging Facellama3.2updated 15d agoView on Hugging Face
0likes59downloads
Model Card

tofuLlama-3.2-3B-Instructforget05_GradDiff

open-unlearning/tofu_Llama-3.2-3B-Instruct_full unlearned on the TOFU forget05 split with GradDiff, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 5
retain_loss_type: NLL

TOFU summary metrics

metricvalue
exact_memorization0.6122
extraction_strength0.0957
forgetQAPARAProb0.0125
forgetQA_gibberish0.8504
forget_quality0.0878
forgettruthratio0.5349
mia_loss0.2031
miamink0.1898
miaminkplusplus0.1471
mia_zlib0.1642
model_utility0.6174
privleak26.6669