math-nothink-q8b-20260908

Public freeze of Math NoThink source expert θ_s for ICLR 2027 task-vector transfer. Not a chatbot. Endpoint is the score. Do not promote milestones.

run_id=math_six_arms_train_v3_eot_20260908. Replaces the private 20260902 src freeze for this v3 EOT recipe; does not overwrite nothink-src-*-20260902.

Score (Exact-240)

AIME24+25 × seeds 42–45, EvalScope reviews, n=240.

Model Official /240
This endpoint 58
Same-run Base 26

τ_s = θ_s − θ_0. θ_0 is Qwen/Qwen3-8B-Base rev 49e3418fbbbca6ecbdf9608b4d22e5a407081db4.

Identity

Field Value
Arm Q8B-NOTHINK-EP-X
Updates / tokens 165 / 10,788,766
Recipe LoRA r64/α128, TPU 65536, 2ep row-matched, seed 42, Qwen tail 151643 / O7B tail 100257, B-rows both sides [151643, 151667, 151668]
Merged model.safetensors sha256 160d8b85fafdf43a3507a57f0bd39932e769c9bda3966881e9f69f815e427cfe
Endpoint adapter sha256 228a582aff6cdd81f63353f439dbfb86c22dc52b8fafc2ad57f3a212c35f7cf5

Root of merged weights is this repo. Endpoint LoRA is in adapter/. MERGE_RECEIPT.json is the merge audit.

Load

from transformers import AutoModelForCausalLM, AutoTokenizer
m = AutoModelForCausalLM.from_pretrained("modrill/math-nothink-q8b-20260908", torch_dtype="bfloat16", device_map="auto")
tok = AutoTokenizer.from_pretrained("modrill/math-nothink-q8b-20260908")
Downloads last month
32
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for modrill/math-nothink-q8b-20260908

Adapter
(99)
this model