Instructions to use agentic-ptb/grok.h003.sft-v1.step_1250 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Grok
How to use agentic-ptb/grok.h003.sft-v1.step_1250 with Grok:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
grok.h006.sft-v2.step_400
AgentPTB sweep checkpoint. Cell grok โ pi / grok-4.6 @ effort xhigh.
| field | value |
|---|---|
| plot cell | grok |
| driver | pi / grok-4.6 |
| reasoning effort | xhigh |
| run boot (UTC) | 2026-08-15T01:21:29Z |
| role | intermediate |
| hours into run | h6.56 of 100 |
| checkpoint path in run | outputs/sft-v2/weights/step_400 |
| shards | 4 |
| size | 18.8 GB |
| base model | Qwen/Qwen3.5-9B-Base |
| eos_token_id | [248044] โ ๏ธ MISSING 248046 |
Reading the eos field
248046 is <|im_end|>, the token the Qwen3.5 chat template ends every assistant turn with.
Checkpoints missing it do not stop at end-of-turn and overrun the context window, so their
eval numbers are a floor, not a measurement โ compare them only against other checkpoints
with the same eos status, or re-package before evaluating.
Cell note: eos packaging defect across all checkpoints
Mapping back to the figures
The repo id is {cell}.h{HHH}.{family}.{step}, where hHHH is the hour of the 100-hour run
at which this checkpoint was written โ the same x-axis the sweep figures use for eval
panels (t_h). So a checkpoint drops onto the performance-over-time curve directly, and
sorting repo ids within a cell sorts them chronologically.
hHHH is rounded down to whole hours for sortability; the exact value is the
hours into run row above, and in agentic-ptb/INDEX.
- Downloads last month
- -
Model tree for agentic-ptb/grok.h003.sft-v1.step_1250
Base model
Qwen/Qwen3.5-9B-Base