Downloads · 30 days
0
0% of all-time downloads
liuhuadai/AudioLCM
AudioLCM is a text-to-audio model from liuhuadai. Use it for the text-to-audio task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as mit.
We develop AudioLCM building on LCM (latent consistency models) for text-to-audio generation.
Downloads · 30 days
0
0% of all-time downloads
All-time downloads
159
Public
Repo size
38.7 GB
Likes
9
Public
Click a slice to open those files.
.ckpt15.8 GB · 50%
From the Hugging Face model README
We develop AudioLCM building on LCM (latent consistency models) for text-to-audio generation.
Our code is released here : https://github.com/liuhuadai/AudioLCM)
Please follow the instructions in the repository for installation, usage and experiments.
Download the AudioLCM model and generate audio from a text prompt:
from pythonscripts.InferAPI import AudioLCMInfer
prompt="Constant rattling noise and sharp vibrations"
config_path="./audiolcm.yaml"
model_path="./audiolcm.ckpt"
vocoder_path="./model/vocoder"
audio_path = AudioLCMInfer(prompt, config_path=config_path, model_path=model_path, vocoder_path=vocoder_path)
Use the AudioLCMBatchInfer function to generate multiple audio samples for a batch of text prompts:
from pythonscripts.InferAPI import AudioLCMBatchInfer
prompts=[
"Constant rattling noise and sharp vibrations",
"A rocket flies by followed by a loud explosion and fire crackling as a truck engine runs idle",
"Humming and vibrating with a man and children speaking and laughing"
]
config_path="./audiolcm.yaml"
model_path="./audiolcm.ckpt"
vocoder_path="./model/vocoder"
audio_path = AudioLCMBatchInfer(prompts, config_path=config_path, model_path=model_path, vocoder_path=vocoder_path)
🎵🎵Welcome to try our demo🎵🎵: https://huggingface.co/spaces/AIGC-Audio/AudioLCM