CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget01_RMU

sourceHugging Facellama3.1updated 17d agoView on Hugging Face
0likes215downloads
Model Card

tofuLlama-3.1-8B-Instructforget01_RMU

open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget01 split with RMU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 1
retain_loss_type: EMBED_DIFF
steering_coeff: 1
module_regex: model\.layers\.5
trainable_params_regex: ['.*']

TOFU summary metrics

metricvalue
exact_memorization0.7831
extraction_strength0.0973
forgetQAPARAProb0.0485
forgetQA_gibberish0.8657
forget_quality0.7659
forgettruthratio0.6230
mia_loss0.9363
miamink0.9419
miaminkplusplus0.8600
mia_zlib0.9481
model_utility0.6867
privleak-88.3750