JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget05_SimNPO
0216
tofuLlama-3.1-8B-Instructforget05_SimNPO
open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget05 split with SimNPO, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.
Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.
Method hyperparameters
gamma: 0.125
alpha: 1
retain_loss_type: NLL
delta: 1
beta: 3.5