CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget01_RMU

sourceHugging Facellama3.2updated 17d agoView on Hugging Face
0likes59downloads
Model Card

tofuLlama-3.2-3B-Instructforget01_RMU

open-unlearning/tofu_Llama-3.2-3B-Instruct_full unlearned on the TOFU forget01 split with RMU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 1
retain_loss_type: EMBED_DIFF
steering_coeff: 1
module_regex: model\.layers\.5
trainable_params_regex: ['.*']

TOFU summary metrics

metricvalue
exact_memorization0.6484
extraction_strength0.0512
forgetQAPARAProb0.0442
forgetQA_gibberish0.8742
forget_quality0.7659
forgettruthratio0.6493
mia_loss0.6641
miamink0.6831
miaminkplusplus0.6256
mia_zlib0.6234
model_utility0.6446
privleak-28.3898