Downloads · 30 days
38
47% of all-time downloads
nightscape/Intern-S2-Mobius-4bit-mlx-mtp
Intern-S2-Mobius-4bit-mlx-mtp is a text generation model from nightscape. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as apache-2.0.
Companion repo — not a standalone model. This holds the Multi-Token-Prediction (MTP) head for nightscape/Intern-S2-Mobius-4bit-mlx, the base 4-bit MLX conversion of internlm/Intern-S2-Mobius.
Downloads · 30 days
38
47% of all-time downloads
All-time downloads
81
Public
Repo size
475 MB
Likes
0
Public
Click a slice to open those files.
.safetensors475 MB · 100%
From the Hugging Face model README
Companion repo — not a standalone model. This holds the Multi-Token-Prediction (MTP) head for
nightscape/Intern-S2-Mobius-4bit-mlx, the
base 4-bit MLX conversion of internlm/Intern-S2-Mobius.
Files:
model-mtp-head.safetensors — the MTP head weights (~463 MB).config.json — the checkpoint config with text_config.mtp_num_hidden_layers: 1 and the mtp-layer
quantization entries (mtp gate at 8-bit).The base model's 5 weight shards, tokenizer, and remote model code are not duplicated here — download the base repo and overlay this head + config alongside them.
omlx loads an MTP-enabled model as base weights plus this head:
# put the base repo on disk, then drop the head + config next to it:
cp model-mtp-head.safetensors config.json <base-dir>/
Then serve with the omlx build that ships the interns2_mobius MTP driver (nightscape/omlx,
add-interns2-mobius).
omlx interns2_mobius
MTP driver. Stock mlx-lm does not load this head.--trust-remote-code is required (custom InternS1 tokenizer + interns2_mobius code).Apache-2.0 — see LICENSE. Model by InternLM; MTP head weights from the internlm/Intern-S2-Mobius
checkpoint, conversion and MLX port by this repo's authors.