Llemma-metamath-34b / README.md

Update README.md

ae63ee4 verified 2 months ago

359 Bytes

metadata

license: apache-2.0

This model is Llemma-34b model used in the paper "An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models". It's based on Llemma-34b and was further finetuned MetaMath with special format for reward. Each step starts with "Step" and ends with "\u043a\u0438".