Downloads · 30 days
0
Dr-Loser/AV-GRPO
AV-GRPO is a text-to-video model from Dr-Loser. Use it when you need video from a text prompt. The card lists the license as other.
AV-GRPO is a modality-anchored online diffusion reinforcement learning framework for joint audio-video generation. It enables full-parameter or LoRA training of the 22B LTX-2.3 model on 8 A800 GPUs.
Downloads · 30 days
0
Access
Public
Updated Sep 25, 2026
Repo size
46.1 GB
Likes
0
Public
Click a slice to open those files.
.safetensors46.1 GB · 100%
From the Hugging Face model README
AV-GRPO is a modality-anchored online diffusion reinforcement learning framework for joint audio-video generation. It enables full-parameter or LoRA training of the 22B LTX-2.3 model on 8 A800 GPUs.
This repository contains the model introduced in AV-GRPO: Modality-Anchored Decoupling Diffusion Reinforcement Learning for Joint Audio-Video Generation.
For training, inference, and usage details, please refer to the GitHub repository.