Downloads · 30 days
0
ChrisColeTech/Anima-turbo
Anima-turbo is a text-to-image model from ChrisColeTech. Use it when you need an image from a text prompt. The card lists the license as unknown.
A 4B anime-native turbo model — ≈6 s per 1024² image on an RTX 5090 at 10 steps, with a ControlNet-LLLite inpainting patch for masked edits.
Downloads · 30 days
0
Access
Public
Updated Aug 3, 2026
Repo size
5.7 GB
Likes
0
Public
Click a slice to open those files.
.safetensors5.7 GB · 100%
From the Hugging Face model README
A 4B anime-native turbo model — ≈6 s per 1024² image on an RTX 5090 at 10 steps, with a ControlNet-LLLite inpainting patch for masked edits.
What this repo is: the Anima Turbo denoiser, the LLLite inpainting patch, a Qwen3-0.6B text encoder and the Qwen-Image VAE — weights only, not a retrain. The settings below are the values these weights are actually run with day to day.
Anime is what this model is for — these are 10-step, guidance 1.0 renders.
<table> <tr><td><img src="https://huggingface.co/ChrisColeTech/Anima-turbo/resolve/main/samples/anime-rooftop.png" width="380" alt="anime girl on a rooftop at sunset"></td><td><img src="https://huggingface.co/ChrisColeTech/Anima-turbo/resolve/main/samples/anime-knight.png" width="380" alt="anime silver-armored knight in a ruined cathedral"></td></tr> <tr><td><sub>**prompt:** `masterpiece, best quality, anime illustration of a girl in a school uniform standing on a rooftop at sunset, detailed cel shading, clean linework, dramatic sky` — 1024×1024, 10 steps, guidance 1.0, seed 21</sub></td><td><sub>**prompt:** `masterpiece, best quality, anime illustration, a silver-armored knight with a glowing blue sword in a ruined cathedral, dramatic lighting, detailed` — 1024×1024, 10 steps, guidance 1.0, seed 7</sub></td></tr> <tr><td><img src="https://huggingface.co/ChrisColeTech/Anima-turbo/resolve/main/samples/anime-street.png" width="380" alt="anime background art of a japanese street in summer"></td><td><img src="https://huggingface.co/ChrisColeTech/Anima-turbo/resolve/main/samples/inpaint-source.png" width="380" alt="anime girl in a white dress in a flower meadow"></td></tr> <tr><td><sub>**prompt:** `masterpiece, best quality, anime background art of a quiet japanese street in summer, blue sky, detailed clouds, vibrant colors, no people` — 1024×1024, 10 steps, guidance 1.0, seed 22</sub></td><td><sub>**prompt:** `masterpiece, best quality, anime illustration, a young woman with long black hair in a white summer dress standing in a flower meadow, blue sky` — 1024×1024, 10 steps, guidance 1.0, seed 21</sub></td></tr> </table>img2img + mask)img2img on this model is masked inpainting through the LLLite patch, not
strength-based restyling. Supply the source image, a white-on-black mask of the
region to change, and a full prompt describing the desired result.
wearing a wide straw sun hat on the
same mask produced no hat, only a slight expression change. LLLite is a
control patch conditioning an existing composition, not an instruction
editor — write the whole scene as you want it, and expect changes in the
masked region to stay close to the underlying shapes.img2img is effectively a no-op — the control signal
reproduces the input. Style-transfer prompts return the source image
unchanged. If nothing is changing, check that a mask was actually sent.Values this build is run with in practice.
| Parameter | Production value | Meaning |
|---|---|---|
width | 1024 | Output width in pixels |
height | 1024 | Output height in pixels |
steps | 10 | Denoising steps |
guidance | 1.0 | Distilled — CFG is not used |
shift | 3.0 | Flow-matching timestep shift |
strength | 1.0 | LLLite control strength (not denoise strength) |
start_percent / end_percent | 0.0 / 1.0 | Fraction of the schedule the control patch is active for |
Supported modes: txt2img, img2img (masked inpainting)
strength here is ControlNet conditioning strength, not img2img denoise
strength — it scales how hard the LLLite patch holds the source, and the
engine also accepts it as control_strength. Lower it to let the masked
region drift further from the original.masterpiece, best quality, anime illustration, …); without them
output drifts toward generic digital painting.studio photograph
yields painterly, illustration-inflected results — use a photo model for
that and this one for anime.Components ship as separate files: the denoiser, the LLLite inpainting patch
under split/model_patches/, the Qwen3-0.6B text encoder, and the VAE. Any
loader that accepts explicit per-component paths and can apply a ControlNet-
LLLite patch can consume this directly.
| File | Size | Role |
|---|---|---|
split/anima-turbo-v1.0.safetensors | 4.18 GB | denoiser (main weights) |
split/model_patches/anima-lllite-inpainting-v2.safetensors | 65.8 MB | ControlNet-LLLite inpainting patch |
split/text_encoders/qwen_3_06b_base.safetensors | 1.19 GB | Qwen3-0.6B text encoder |
split/vae/qwen_image_vae.safetensors | 254 MB | VAE |
unknown in this repo's metadata. Refer to the upstream model's license for redistribution and commercial-use terms.