math_nothink_lora_0.1

Localize-and-Stitch merge of a math nothink SFT into Qwen/Qwen3-4B-Base.

Lineage

Recipe

Task vector (\tau = W_{\text{SFT}} - W_{\text{base}}), then a binary mask is trained on MergeBench/math_val (64-shot) and the masked vector is stitched back into the base.

Setting Value
Task math
Sparsity 0.1
Mask LR 1e7
Mask epochs 10
L1 1e-5
Mask train SGD, batch 1, grad_accum 2, seq ≈ 2048, sdpa
Downloads last month
30
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Montalte/math_nothink_lora_0.1

Finetuned
(378)
this model