SAM-AI-Reasoning-14B

Official submission of SAM-AI Reasoning Engine trained with Group Relative Policy Optimization (GRPO).

Evaluation

Submitted for independent verification on the Hugging Face Open LLM Leaderboard v2:

  • IFEval (Instruction Following)
  • BBH (Big Bench Hard)
  • MATH-Hard (Advanced Competition Mathematics)
  • GPQA Diamond (PhD-level Science)
  • MuSR (Multi-step Soft Reasoning)
  • MMLU-Pro (Professional Knowledge)
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Samrish2009/SAM-AI-Reasoning-14B

Finetuned
(97)
this model