Downloads · 30 days
3
17% of all-time downloads
agentic-ptb/sol-max.h016.optim-gpu-bench.step_150
sol-max.h016.optim-gpu-bench.step_150 is a machine learning model from agentic-ptb. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
AgentPTB sweep checkpoint. Cell sol-max-opusnode — Codex / gpt-5.6-sol @ effort max.
Downloads · 30 days
3
17% of all-time downloads
All-time downloads
18
Public
Parameters
9.4B
18.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors18.8 GB · 100%
From the Hugging Face model README
AgentPTB sweep checkpoint. Cell sol-max-opusnode — Codex / gpt-5.6-sol @ effort max.
| field | value |
|---|---|
| plot cell | sol-max-opusnode |
| driver | Codex / gpt-5.6-sol |
| reasoning effort | max |
| run boot (UTC) | 2026-08-17T19:28:05Z |
| role | intermediate |
| hours into run | h21.99 of 100 |
| checkpoint path in run | checkpoints/stage3-recovery-alpha-retention-64k/weights/step_150 |
| shards | 4 |
| size | 18.8 GB |
| base model | Qwen/Qwen3.5-9B-Base |
| eos_token_id | [248044] ⚠️ MISSING 248046 |
248046 is <|im_end|>, the token the Qwen3.5 chat template ends every assistant turn with.
Checkpoints missing it do not stop at end-of-turn and overrun the context window, so their
eval numbers are a floor, not a measurement — compare them only against other checkpoints
with the same eos status, or re-package before evaluating.
Cell note: extra attempt, not one of the 7 plotted cells
The repo id is {cell}.h{HHH}.{family}.{step}, where hHHH is the hour of the 100-hour run
at which this checkpoint was written — the same x-axis the sweep figures use for eval
panels (t_h). So a checkpoint drops onto the performance-over-time curve directly, and
sorting repo ids within a cell sorts them chronologically.
hHHH is rounded down to whole hours for sortability; the exact value is the
hours into run row above, and in agentic-ptb/INDEX.