Downloads · 30 days
0
ipfipfipf/Qwen3.5-4B-MathCodeSearch-SDPO-E-Step30
Qwen3.5-4B-MathCodeSearch-SDPO-E-Step30 is a text generation model from ipfipfipf. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
This repository contains the complete raw Megatron Core distributed checkpoint from the Qwen3.5-4B Math+Code+Search SDPO-E run at rollout step 30.
Downloads · 30 days
0
Access
Public
Updated Aug 21, 2026
Repo size
58.9 GB
Likes
0
Public
Click a slice to open those files.
.distcp58.9 GB · 100%
From the Hugging Face model README
This repository contains the complete raw Megatron Core distributed checkpoint from the Qwen3.5-4B Math+Code+Search SDPO-E run at rollout step 30.
jzqlsb95iter_0000029 (rollout step 30; rollout indices are zero-based)torch_dist distributed checkpointThe upload preserves the original checkpoint tree, including all 16 .distcp
shards, .metadata, common.pt, metadata.json,
latest_checkpointed_iteration.txt, and the rollout dataset state files for
steps 10, 20, and 30.
This is not a Transformers from_pretrained checkpoint. Resume or inspect
it with a compatible Megatron Core/MILES-SDPO environment and point the
checkpoint loader at the repository root. The original tensor-parallel layout
and the torch_dist checkpoint format should be retained when resuming.
This is an intermediate checkpoint; the run stopped after step 30 rather than completing the planned 51 rollouts.