b7 stage 3 -- SFR + expanding rank (32->128), capacity-matched to b6

  • LoRA rank / alpha: 96 / 192 (scaling: alpha/r)
  • Full rank schedule: 32 -> 64 -> 96 -> 128 -> 128
  • Replay (cumulative levels): True
  • Stage partition: difficulty
  • Cumulative train examples this stage: 3329
  • Validation split seed: 42 (5% of train, stratified by level; test set never used for selection)
  • Code commit: 1720a936dde4503227fe375f958eda65e36ab8fd
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support