Instructions to use ReasoningRegisters/olmo32b-cue with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use ReasoningRegisters/olmo32b-cue with PEFT:
Task type is invalid.
- Notebooks
- Google Colab
- Kaggle
olmo32b-cue โ GRPO with forced cue opener
LoRA adapters from the cue-forced GRPO run on Olmo-3-1125-32B (MATH training set,
300 steps; every rollout prefilled with the ".\n\nOkay," opener). Companion to the
vanilla run at ReasoningRegisters/olmo32b.
Checkpoints 50-300 (every 50 steps): adapter weights + tokenizer + trainer state.
Training rollout logs: ReasoningRegisters/olmo32b-cue-rollouts (dataset).
- Downloads last month
- -
Model tree for ReasoningRegisters/olmo32b-cue
Base model
allenai/Olmo-3-1125-32B