Downloads · 30 days
5
13% of all-time downloads
ajh-code/Mage-Flow-NVFP4-Balanced-AJH
Mage-Flow-NVFP4-Balanced-AJH is a text-to-image model from ajh-code. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as other.
Mage-Flow-NVFP4-Balanced-AJH is a portable, runnable quantized version of microsoft/Mage-Flow. Standalone Hugging Face Mage-Flow repository with 44 native NVFP4 transformer MLP projections and four late image MLP modu…
Downloads · 30 days
5
13% of all-time downloads
All-time downloads
39
Public
Parameters
3.4B
10.2 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors10.2 GB · 100%
How the weights are stored.
BF162.5B · 72%
From the Hugging Face model README
Mage-Flow-NVFP4-Balanced-AJH is a portable, runnable quantized version of
microsoft/Mage-Flow.
Standalone Hugging Face Mage-Flow repository with 44 native NVFP4 transformer MLP projections and four late image MLP modules kept in BF16.
This repository is intended to remain discoverable as a quantized
microsoft/Mage-Flow derivative while avoiding any runtime dependency on
a separate local models/ checkout. The complete transformer component,
quantized text encoder, VAE, scheduler, vendored inference code, and
native runtime are all packaged inside the repository layout.
Related releases:

Generated directly with this balanced package, without upscaling
or post-processing.
A 4K resolution high detail photo realistic image of the top half of a cyborg woman with dark black hair, striking blue eyes that have a very subtle glow in the iris, standing side profile, head tilted up towards the sky with a questioning expression, she has subtle gaps in her skin that hint at a robotic nature, outdoor forest night setting, sky filled with bright brilliant stars that glow against the dark setting, nebula visible
Settings: 1280×1280, 20 steps, CFG 5, static shift 6, seed
3334072683.
444amax10.75BF16 passthrough modules:
transformer_blocks.9.img_mlp.net.2transformer_blocks.10.img_mlp.net.2transformer_blocks.11.img_mlp.net.0.projtransformer_blocks.11.img_mlp.net.2The Qwen3-VL text encoder uses the same mixed policy as the Fast release: 224 NVFP4 projections in blocks 2–33, 14 FP8 projections in blocks 1 and 34, with blocks 0 and 35 plus embeddings, norms, biases, and the vision tower retained in BF16.
The tested stack is Linux x86-64, NVIDIA Blackwell SM120, CUDA 13.1,
Python 3.11, PyTorch 2.13.0+cu130, comfy-kitchen==0.2.22, and
flash-attn==2.8.3. Install into a virtual environment using the
included requirements.txt; do not install these packages system-wide.
python3.11 -m venv .venv
source .venv/bin/activate
python -m pip install --upgrade pip
python -m pip install -r requirements.txt
CUDA_HOME=/usr/local/cuda-13.1 python -m pip install --no-build-isolation flash-attn==2.8.3
CUDA_VISIBLE_DEVICES=0 .venv/bin/python generate.py \
--model ajh-code/Mage-Flow-NVFP4-Balanced-AJH \
--prompt 'A detailed watercolor fox reading under an old oak tree' \
--output fox.png --height 1024 --width 1024 --steps 20 --seed 1
The included binaries target the tested stack. Run build_native.sh
after changing PyTorch, CUDA, or the C++ ABI.
The controlled matched benchmark measured the Balanced transformer at
about 1.43x BF16 throughput (30% less generation time), and the
Quality transformer at about 1.22x BF16 throughput (18% less
generation time). Projected transformer plus text-checkpoint storage
is about 9.89 GB for Balanced and 10.97 GB for Quality, versus
17.12 GB for BF16. These are policy-level measurements from the
research suite, not universal hardware guarantees.
mage_flow_nvfp4_* transformer modules.MAGE_NVFP4_* environment variables.python validate_release.py
Mage-Flow and the vendored Mage inference source are Copyright (c) 2026 Microsoft and MIT licensed. Qwen3-VL and the mixed NVFP4/FP8 text checkpoint are Apache-2.0 licensed. See the included license files and third-party notices for details.