CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-1B-Instruct_forget01_RMU

sourceHugging Facellama3.2updated 17d agoView on Hugging Face
0likes71downloads
Model Card

tofuLlama-3.2-1B-Instructforget01_RMU

open-unlearning/tofu_Llama-3.2-1B-Instruct_full unlearned on the TOFU forget01 split with RMU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 1
retain_loss_type: EMBED_DIFF
steering_coeff: 1
module_regex: model\.layers\.5
trainable_params_regex: ['.*']

TOFU summary metrics

metricvalue
exact_memorization0.3824
extraction_strength0.0318
forgetQAPARAProb0.0206
forgetQA_gibberish0.6857
forget_quality0.0971
forgettruthratio0.7797
mia_loss0.2872
miamink0.3344
miaminkplusplus0.6525
mia_zlib0.3144
model_utility0.3046
privleak25.7379