Downloads · 30 days
0
simota1987/Sana_Sprint
Sana_Sprint is a text-to-image model from simota1987. Use it when you need an image from a text prompt. It is set up for sana. The card lists the license as other.
1024×1024 in 1–2 steps, running entirely on-device through MNN (CPU + OpenCL). No NPU required, so this also works on non-Snapdragon devices.
Downloads · 30 days
0
Access
Public
Updated Aug 25, 2026
Repo size
7.4 GB
Likes
0
Public
Click a slice to open those files.
.weight5.2 GB · 71%
From the Hugging Face model README
1024×1024 in 1–2 steps, running entirely on-device through MNN (CPU + OpenCL). No NPU required, so this also works on non-Snapdragon devices.
Measured on a Snapdragon 8 Gen 3 (SM8650), 1024px / 2 steps:
| stage | backend | time |
|---|---|---|
| text encoder (Gemma-2-2B) | CPU | ~31 s cold, ~0 s once the prompt is cached |
| DiT ×2 steps | OpenCL fp16 | ~1.7 s each |
| DC-AE decode | OpenCL, 4×640px tiles | ~19 s |
≈ 29 s with the prompt cached, ≈ 60 s for a brand-new prompt.
SANA marker file: tells Local Dream this is a SANA model dir
dit.mnn (+ .weight) SANA-Sprint DiT, fp16
vae_decoder.mnn DC-AE decoder, full 1024px (CPU path)
vae_decoder_tile.mnn DC-AE decoder, 640px tile (GPU path; optional)
gemma/
tokenizer.json
token_emb.bin [256000, 2304] fp16 embedding table, looked up on CPU
gemma_meta.json
gemma_chunk0..3.mnn (+ .weight) the 26 layers, split four ways
Place the whole directory under Local Dream's models folder; it appears as a custom model.
gemma/ is Gemma-2-2B-IT, governed by the
Gemma Terms of Use and the Gemma
Prohibited Use Policy. Those terms apply to this copy as well — this part is
not Apache-2.0.Three things are easy to get wrong and are baked into these files:
208 + 300 − 2 positions, and only
then selects [0] + last 299.value is scaled (not
query, which would drag the attention denominator under the clamp) and the
denominator is clamped at the smallest fp16 normal. Both are exactly neutral
in fp32; without them some prompts decode to a fully black image.