Downloads · 30 days
77
9% of all-time downloads
EricRollei/HunyuanImage-3.0-Instruct-INT8-v2
HunyuanImage-3.0-Instruct-INT8-v2 is a text-to-image model from EricRollei. Use it when you need an image from a text prompt. It is set up for transformers. The card lists the license as other.
INT8 quantization of the HunyuanImage-3.0 Instruct model (v2). Supports text-to-image, image editing, multi-image fusion, and Chain-of-Thought prompt enhancement (recaption/thinkrecaption).
Downloads · 30 days
77
9% of all-time downloads
All-time downloads
858
Public
Parameters
83B
88.8 GB on disk
Likes
3
Public
Click a slice to open those files.
.safetensors88.8 GB · 100%
How the weights are stored.
I877.3B · 93%
From the Hugging Face model README
INT8 quantization of the HunyuanImage-3.0 Instruct model (v2). Supports text-to-image, image editing, multi-image fusion, and Chain-of-Thought prompt enhancement (recaption/think_recaption).
v2 uses improved quantization with more precise skip-module selection, keeping attention projections and critical embedding layers in full BF16 precision for better image quality.
think_recaption mode for highest quality| Component | Memory |
|---|---|
| Weight Loading | ~80 GB weights |
| Inference (additional) | ~12-20 GB inference |
| Total | ~92-100 GB |
Recommended Hardware:
Layers quantized to INT8:
Kept in full precision (BF16):
This model is designed to work with the Comfy_HunyuanImage3 custom nodes:
cd ComfyUI/custom_nodes
git clone https://github.com/EricRollei/Comfy_HunyuanImage3
int8 precisionThe Instruct model supports three generation modes:
| Mode | Description | Speed |
|---|---|---|
image | Direct text-to-image, prompt used as-is | Fastest |
recaption | Model rewrites prompt into detailed description, then generates | Medium |
think_recaption | CoT reasoning -> prompt enhancement -> generation (best quality) | Slowest |
Block swap allows running INT8 and BF16 models on GPUs with less VRAM than the full model requires. The system keeps N transformer blocks on CPU and swaps them to GPU on demand during each diffusion step.
| blocks_to_swap | VRAM Saved | Recommended For |
|---|---|---|
| 0 | 0 GB | 96GB+ GPU (no swap needed) |
| 4 | ~10 GB | 80-90GB GPU |
| 8 | ~20 GB | 64-80GB GPU |
| 16 | ~40 GB | 48-64GB GPU |
| -1 (auto) | varies | Let the system decide |
This is a quantized derivative of Tencent's HunyuanImage-3.0 Instruct.
This model inherits the license from the original Hunyuan Image 3.0 model: Tencent Hunyuan Community License