Downloads · 30 days
26
16% of all-time downloads
byungjun-kim/DWM-CogVideoX-Fun-5b-LoRA
DWM-CogVideoX-Fun-5b-LoRA is a image-to-video model from byungjun-kim. Use it for the image-to-video task on the model card, and read the license before you ship it in a product. It is set up for diffusers. The card lists the license as other.
This repository contains a checkpoint-8000 release of the DWM CogVideoX static-scene + hand-video concat model.
Downloads · 30 days
26
16% of all-time downloads
All-time downloads
166
Public
Repo size
133 MB
Likes
0
Public
Click a slice to open those files.
.safetensors133 MB · 100%
From the Hugging Face model README
This repository contains a checkpoint-8000 release of the DWM CogVideoX static-scene + hand-video concat model.
It was trained on top of:
alibaba-pai/CogVideoX-Fun-V1.1-5b-InPThis checkpoint is a derivative of CogVideoX-Fun-V1.1-5b-InP, and its use must comply with the upstream CogVideoX license.
The checkpoint includes:
pytorch_lora_weights.safetensors: LoRA weightsnon_lora_weights.safetensors: non-LoRA projection weights for patch_embed.projdwm_cogvideox_5b_lora.yaml: training config used for this checkpointThis checkpoint is intended to be used with the dwm repository inference script:
python training/cogvideox/inference.py \
--checkpoint_path /path/to/this_repo \
--experiment_config /path/to/this_repo/dwm_cogvideox_5b_lora.yaml \
--data_root /path/to/data \
--video physics/videos/00001.mp4 \
--output_dir outputs_infer/dwm_cogvideox_hf
The expected sibling inputs are:
videos/<stem>.mp4videos_static/<stem>.mp4videos_hands/<stem>.mp4prompts_rewrite/<stem>.txtnon_lora_weights.safetensors was converted from the original non_lora_weights.pt checkpoint artifact and stores:
patch_embed.proj.weightpatch_embed.proj.biasCogVideoX-Fun-V1.1-5b-InP.