staged stage 3 -- Curriculum: L1 -> L1,2 -> ... -> L1..5 (constant rank)
- LoRA rank / alpha: 32 / 64 (scaling: alpha/r)
- Full rank schedule: 32 -> 32 -> 32 -> 32 -> 32
- Replay (cumulative levels): True
- Stage partition: difficulty
- Early stopping patience: 1000000
- Cumulative train examples this stage: 3329
- Validation split seed: 42 (5% of train, stratified by level; test set never used for selection)
- Code commit: 1967848696478b02e365550eb6da186f2d5b2bcf
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support