Add model card
#1
by nielsr HF Staff - opened
This PR adds a model card for the model presented in Learn to Reason Efficiently with Adaptive Length-based Reward Shaping.
This PR adds a model card for the model presented in Learn to Reason Efficiently with Adaptive Length-based Reward Shaping.