Downloads · 30 days
3.6K
43% of all-time downloads
realrebelai/LingBot_ComfyUI
LingBot_ComfyUI is a image-to-video model from realrebelai. Use it for the image-to-video task on the model card, and read the license before you ship it in a product. It is set up for diffusers. The card lists the license as other.
<div align="center" <video src="https://cdn-uploads.huggingface.co/production/uploads/68761990332d15464ccc8dee/4c5xvrkNvb1Nde1xUK1Eu.mp4" width="400" height="400" autoplay loop muted playsinline </video </div
Downloads · 30 days
3.6K
43% of all-time downloads
All-time downloads
8.4K
Public
Repo size
12.2 GB
Likes
11
Public
Click a slice to open those files.
.safetensors12.2 GB · 100%
From the Hugging Face model README
Repackaged weights of robbyant/lingbot-video-dense-1.3b for use with the ComfyUI_Rebels_LingBot node pack. No weight modification — the original diffusers shards are merged/renamed into single files for ComfyUI's model folders. All configs ship inside the node pack.
Runs T2V and TI2V on an RTX 3070 8GB / 16GB RAM: the Qwen3-VL encoder is loaded on CPU, used, and freed before the 1.3B DiT (2.79GB bf16, fully VRAM-resident) starts.
| File | Put in | Size |
|---|---|---|
LingBot_1.3b_DiT.safetensors | ComfyUI/models/diffusion_models/ | ~2.8 GB |
LingBot_text-encoder.safetensors | ComfyUI/models/text_encoders/ | ~8 GB |
LingBot_vae.safetensors | ComfyUI/models/vae/ | ~0.5 GB |
The text encoder is the repo's Qwen3-VL, shards merged into one file (bit-identical
tensors). lm_head.weight is intentionally absent — it is tied to embed_tokens; the
node pack re-ties it at load.
~14 s/it on a 3070 at the settings above (sequential CFG); ≈10 minutes per 81-frame clip.
Weights inherit the upstream LingBot-Video license (see license_link). This repo is
a format repack only; review the upstream terms before commercial use or redistribution.
Repack + nodes by RealRebelAI.