Downloads · 30 days
66
8% of all-time downloads
frankleeeee/CausalForcing-Wan2.1-T2V-1.3B-Diffusers
CausalForcing-Wan2.1-T2V-1.3B-Diffusers is a text-to-video model from frankleeeee. Use it when you need video from a text prompt. It is set up for diffusers. The card lists the license as apache-2.0.
Diffusers-layout conversion of the chunk-wise Causal Forcing generator (zhuhz22/Causal-Forcing, chunkwise/causalforcing.pt, generator weights) from thu-ml/Causal-Forcing — "Causal Forcing: Autoregressive Diffusion Dis…
Downloads · 30 days
66
8% of all-time downloads
All-time downloads
788
Public
Parameters
1.4B
26.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors26.1 GB · 100%
From the Hugging Face model README
Diffusers-layout conversion of the chunk-wise Causal Forcing generator
(zhuhz22/Causal-Forcing, chunkwise/causal_forcing.pt, generator weights)
from thu-ml/Causal-Forcing —
"Causal Forcing: Autoregressive Diffusion Distillation Done Right" (ICML 2026).
Wan2.1-T2V-1.3B backbone, block-causal DMD student: 3-latent-frame chunks,
21-latent-frame sliding attention window, 4 warped denoising steps
(1000/750/500/250, shift 5.0), no CFG. Non-transformer components are copied
from Wan-AI/Wan2.1-T2V-1.3B-Diffusers.
Converted with
sglang.multimodal_gen.tools.convert_forcing_to_diffusers --preset causal-forcing-chunkwise
for use with the SGLang diffusion runtime:
sglang generate --model-path frankleeeee/CausalForcing-Wan2.1-T2V-1.3B-Diffusers \
--prompt "A stylish woman walks down a Tokyo street..." \
--width 832 --height 480 --num-frames 81 --save-output
Note: the upstream project releases only Wan 1.3B generators; the
Wan2.1-T2V-14B named in its configs is the DMD teacher (real_name), not a
released generator.