HETU-Qwen3-4B-MathReasoning-CotGen

Part of the HETU (Hints Enable True Understanding) model suite.

  • Base model: Qwen/Qwen3-4B
  • Task: math reasoning (AIME, GSM8K, MATH-500, Omni-MATH, GPQA-Diamond, MMLU)
  • Method: trained to generate a full chain-of-thought before producing its output (CotGen)

This is the full merged model (base weights + LoRA adapter merged in), final training checkpoint, in bf16.

See the HETU paper for the training setup, evaluation methodology, and results tables.

Downloads last month
450
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AdarshSingh7647/HETU-Qwen3-4B-MathReasoning-CotGen

Finetuned
Qwen/Qwen3-4B
Finetuned
(1110)
this model

Collection including AdarshSingh7647/HETU-Qwen3-4B-MathReasoning-CotGen