πŸ“Š Benchmark Results

Benchmark Metric Score
ARC-Challenge acc_norm 39.25%
GSM8K exact_match 37.45%
MMLU acc 53.75%

Evaluated using lm-evaluation-harness.

Note: v2 of this model coming soon with stronger reasoning and agentic capabilities. Will fix the "identity confusion" Users may experience with this current model.

Downloads last month
176
Safetensors
Model size
1B params
Tensor type
F16
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for Healshsj/MiniCPM5-1B-Reasoning-Agent-Ultra

Merges
1 model
Quantizations
2 models