tofu_Llama-3.2-3B-Instruct_forget05_RMU

open-unlearning/tofu_Llama-3.2-3B-Instruct_full unlearned on the TOFU forget05 split with RMU, trained with the open-unlearning framework. Used as a weight-unlearning baseline / draft model in the Speculative-Decoding-Unlearning project.

Full training config: .hydra/config.yaml. TOFU evaluation outputs: evals/.

Method hyperparameters

gamma: 1.0
alpha: 1
retain_loss_type: EMBED_DIFF
steering_coeff: 1
module_regex: model\.layers\.5
trainable_params_regex: ['.*']

TOFU summary metrics

metric value
exact_memorization 0.1808
extraction_strength 0.0329
forget_Q_A_PARA_Prob 0.0031
forget_Q_A_gibberish 0.4621
forget_quality 0.0000
forget_truth_ratio 0.7175
mia_loss 0.0450
mia_min_k 0.0617
mia_min_k_plus_plus 0.7944
mia_zlib 0.0346
model_utility 0.6672
privleak 46.6896
Downloads last month
66
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget05_RMU

Finetuned
(31)
this model

Dataset used to train JoaoBoer/tofu_Llama-3.2-3B-Instruct_forget05_RMU