b8 stage 2 -- SFR + constant rank r=102, capacity-matched to b6
- LoRA rank / alpha: 102 / 204 (scaling: alpha/r)
- Full rank schedule: 102 -> 102 -> 102 -> 102 -> 102
- Replay (cumulative levels): True
- Stage partition: difficulty
- Early stopping patience: 2
- Cumulative train examples this stage: 1817
- Validation split seed: 42 (5% of train, stratified by level; test set never used for selection)
- Code commit: c925de05f810a41b16d469627f37f87c9283d7ac
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support