HETU-Qwen3-8B-MathReasoning-CotCond

Part of the HETU (Hints Enable True Understanding) model suite.

  • Base model: Qwen/Qwen3-8B
  • Task: math reasoning (AIME, GSM8K, MATH-500, Omni-MATH, GPQA-Diamond, MMLU)
  • Method: trained on a compact conditioning signal instead of a full generated chain-of-thought (CotCond, HETU's method)

This is the full merged model (base weights + LoRA adapter merged in), final training checkpoint, in bf16.

See the HETU paper for the training setup, evaluation methodology, and results tables.

Downloads last month
-
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AdarshSingh7647/HETU-Qwen3-8B-MathReasoning-CotCond

Finetuned
Qwen/Qwen3-8B
Finetuned
(1993)
this model

Collection including AdarshSingh7647/HETU-Qwen3-8B-MathReasoning-CotCond