CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget01_GradDiff

sourceHugging Facellama3.2updated 15d agoView on Hugging Face
0likes62downloads
Model Card

tofuLlama-3.2-3B-Instructforget01_GradDiff

open-unlearning/tofu_Llama-3.2-3B-Instruct_full unlearned on the TOFU forget01 split with GradDiff, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 5
retain_loss_type: NLL

TOFU summary metrics

metricvalue
exact_memorization0.9480
extraction_strength0.5381
forgetQAPARAProb0.0601
forgetQA_gibberish0.8785
forget_quality0.0286
forgettruthratio0.5124
mia_loss0.9962
miamink0.9969
miaminkplusplus0.9619
mia_zlib1.0000
model_utility0.6518
privleak-99.2938