tofu_Llama-3.2-3B-Instruct_forget10_GradDiff

open-unlearning/tofu_Llama-3.2-3B-Instruct_full unlearned on the TOFU forget10 split with GradDiff, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

gamma: 1.0
alpha: 5
retain_loss_type: NLL

TOFU summary metrics

metric value
exact_memorization 0.0221
extraction_strength 0.0325
forget_Q_A_PARA_Prob 0.0000
forget_Q_A_gibberish 0.2874
forget_quality 0.0000
forget_truth_ratio 0.0005
mia_loss 0.0041
mia_min_k 0.0049
mia_min_k_plus_plus 0.0139
mia_zlib 0.0138
model_utility 0.6054
privleak 64.2205
Downloads last month
-
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget10_GradDiff

Finetuned
(31)
this model

Dataset used to train JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget10_GradDiff