JoaoBoer/tofu_Llama-3.1-8B-Instruct_forget10_GradDiff
0217
tofuLlama-3.1-8B-Instructforget10_GradDiff
open-unlearning/tofu_Llama-3.1-8B-Instruct_full unlearned on the TOFU forget10 split with GradDiff, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.
Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.
Method hyperparameters
gamma: 1.0
alpha: 5
retain_loss_type: NLL