CoolFace
Modelpublic

JoaoBoer/tofu_Llama-3.2-1B-Instruct_forget05_IdkDPO

sourceHugging Facellama3.2updated 17d agoView on Hugging Face
0likes83downloads
Model Card

tofuLlama-3.2-1B-Instructforget05_IdkDPO

open-unlearning/tofu_Llama-3.2-1B-Instruct_full unlearned on the TOFU forget05 split with IdkDPO, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

yaml
gamma: 1.0
alpha: 2
retain_loss_type: NLL
beta: 0.05

TOFU summary metrics

metricvalue
exact_memorization0.6629
extraction_strength0.1092
forgetQAPARAProb0.0909
forgetQA_gibberish0.9683
forget_quality0.0163
forgettruthratio0.6657
mia_loss0.6662
miamink0.6697
miaminkplusplus0.7425
mia_zlib0.5602
model_utility0.0000
privleak-48.1905