CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-1B-Instruct_forget01_PDU

sourceHugging Facellama3.2updated 17d agoView on Hugging Face
0likes76downloads
Model Card

tofuLlama-3.2-1B-Instructforget01_PDU

open-unlearning/tofu_Llama-3.2-1B-Instruct_full unlearned on the TOFU forget01 split with PDU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 100
retain_loss_type: NLL
retain_loss_eps: 0.3
primal_dual: True
dual_step_size: 5
dual_update_upon: step
dual_warmup_epochs: 5
loss_names: ['forget_loss', 'retain_loss']

TOFU summary metrics

metricvalue
exact_memorization0.5829
extraction_strength0.0689
forgetQAPARAProb0.0321
forgetQA_gibberish0.7192
forget_quality0.0286
forgettruthratio0.5861
mia_loss0.3722
miamink0.4500
miaminkplusplus0.9587
mia_zlib0.4416
model_utility0.5849
privleak3.8961