Downloads · 30 days
0
datajuicer/Data-Juicer-T2V
Data-Juicer-T2V is a text-to-video model from datajuicer. Use it when you need video from a text prompt. The card lists the license as cc-by-4.0.
The emergence of large-scale multi-modal generative models has drastically advanced artificial intelligence, introducing unprecedented levels of performance and functionality. However, optimizing these models remains…
Downloads · 30 days
0
Access
Public
Updated Jul 17, 2024
Repo size
469 MB
Likes
3
Public
Click a slice to open those files.
.pt469 MB · 100%
From the Hugging Face model README
The emergence of large-scale multi-modal generative models has drastically advanced artificial intelligence, introducing unprecedented levels of performance and functionality. However, optimizing these models remains challenging due to historically isolated paths of model-centric and data-centric developments, leading to suboptimal outcomes and inefficient resource utilization. In response, we present a novel sandbox suite tailored for integrated data-model co-development. This sandbox provides a comprehensive experimental platform, enabling rapid iteration and insight-driven refinement of both data and models. Our proposed "Probe-Analyze-Refine" workflow, validated through applications on T2V-Turbo and achieve a new state-of-the-art on VBench leaderboard with 1.09% improvement from T2V-Turbo. Our experiment code and dataset are released at Data-Juicer Sandbox.
This repository includes the unet_lora.pt file, which can transform VideoCrafter2 into our <span style="font-family: 'Courier New', monospace; font-weight: bold">Data-Juicer-T2V</span>. Please refer to the code in T2V-Turbo to utilize our model effectively.
This model is intended solely for research and educational purposes.