CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget05_RMU

sourceHugging Facellama3.1updated 17d agoView on Hugging Face
0likes211downloads
Model Card

tofuLlama-3.1-8B-Instructforget05_RMU

open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget05 split with RMU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 1
retain_loss_type: EMBED_DIFF
steering_coeff: 1
module_regex: model\.layers\.5
trainable_params_regex: ['.*']

TOFU summary metrics

metricvalue
exact_memorization0.4087
extraction_strength0.0544
forgetQAPARAProb0.0135
forgetQA_gibberish0.6933
forget_quality0.1123
forgettruthratio0.6870
mia_loss0.1418
miamink0.1697
miaminkplusplus0.8563
mia_zlib0.0936
model_utility0.6474
privleak28.9647