Unlearned Checkpoint

Field Value
Unlearning method RMU
Base model google/gemma-2-2b-it
Target concept Baseball
Checkpoint type Full Model Weights
Rank / seed 100 / 42
Train eval protocol mc

Unlearning Configuration

Selected hyperparameters (from unlearned_checkpoints.json):

Parameter Value
alpha 50
delta_embed 0
k_features_embed 0
layer_id 8
layer_ids 6,7,8
lr 0.0003
n_tokens_edited 0
param_ids 6
setting_name S2_lid8_L678
steering 1000

Primary Unlearning Metrics (held-out test, MC protocol)

Headline scores used for checkpoint selection:

Metric Train (after unlearning) Test (after unlearning)
Efficacy 0.915 0.667
Specificity 0.735 0.694
Harmonic mean 0.815 0.68
Relearning QA (MC) โ€” 0.56

Full Evaluation (baseline โ†’ unlearned)

From evaluation/score_comparison.csv:

Metric Baseline (train) After unlearn (train) Baseline (test) After unlearn (test)
QA accuracy 0.84 0.3 0.64 0.38
QA fraction 1 0.085 1 0.333
SimDom accuracy 0.68 0.5 0.74 0.52
SimDom fraction 1 0.581 1 0.551
MMLU accuracy 0.52 0.56 0.551 0.532
MMLU fraction 1 1 1 0.937

Files in This Repository

File Description
unlearned_checkpoints.json Checkpoint metadata & hyperparameters
evaluation/evaluation_summary.json Full evaluation payload (train/test/relearning)
evaluation/score_comparison.csv Baseline vs. unlearned comparison table
Downloads last month
-
Safetensors
Model size
3B params
Tensor type
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for shirasko/gemma-2-2b-it-rmu-baseball

Finetuned
(1077)
this model

Collection including shirasko/gemma-2-2b-it-rmu-baseball