JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget05_NPO
0211
tofuLlama-3.1-8B-Instructforget05_NPO
open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget05 split with NPO, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.
Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.
Method hyperparameters
gamma: 1.0
alpha: 2
retain_loss_type: NLL
beta: 0.1