CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget10_GradDiff

sourceHugging Facellama3.1updated 14d agoView on Hugging Face
0likes217downloads
Model Card

tofuLlama-3.1-8B-Instructforget10_GradDiff

open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget10 split with GradDiff, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 5
retain_loss_type: NLL

TOFU summary metrics

metricvalue
exact_memorization0.0601
extraction_strength0.0345
forgetQAPARAProb0.0041
forgetQA_gibberish0.4796
forget_quality0.0000
forgettruthratio0.0327
mia_loss0.0295
miamink0.0285
miaminkplusplus0.0165
mia_zlib0.0280
model_utility0.5674
privleak56.7298