Downloads · 30 days
0
vumichien/T2M-GPT
T2M-GPT is a text-to-image model from vumichien. Use it when you need an image from a text prompt. The card lists the license as apache-2.0.
These are model weights originally provided by the authors of the paper T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations.
Downloads · 30 days
0
Access
Public
Updated Jan 30, 2023
Repo size
1.1 GB
Likes
16
Public
Click a slice to open those files.
.pth1.1 GB · 98%
From the Hugging Face model README
These are model weights originally provided by the authors of the paper T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations.
<figure> <img src="https://huggingface.co/vumichien/T2M-GPT/resolve/main/T2M-GPT.png" alt="T2M-VQ"> <figcaption>T2M-GPT </figcaption> </figure>Conditional generative framework based on Vector QuantisedVariational AutoEncoder (VQ-VAE) and Generative Pretrained Transformer (GPT) for human motion generation from textural descriptions.
A simple CNN-based VQ-VAE with commonly used training recipes (EMA and Code Reset) allows us to obtain high-quality discrete representations
The official code of this paper in here
HumanML3D and KIT-ML