staged stage 2 -- Curriculum: L1 -> L1,2 -> ... -> L1..5 (constant rank)

  • LoRA rank / alpha: 32 / 64 (scaling: alpha/r)
  • Full rank schedule: 32 -> 32 -> 32 -> 32 -> 32
  • Replay (cumulative levels): True
  • Stage partition: difficulty
  • Early stopping patience: 1000000
  • Cumulative train examples this stage: 1817
  • Validation split seed: 42 (5% of train, stratified by level; test set never used for selection)
  • Code commit: 1967848696478b02e365550eb6da186f2d5b2bcf
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support