Downloads · 30 days
117
1% of all-time downloads
teticio/audio-diffusion-ddim-256
audio-diffusion-ddim-256 is a machine learning model from teticio. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for diffusers.
De-noising Diffusion Implicit Model trained on teticio/audio-diffusion-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The code to convert from audio to spectrogram and vice versa can…
Downloads · 30 days
117
1% of all-time downloads
All-time downloads
9.4K
Public
Repo size
724 MB
Likes
3
Public
Click a slice to open those files.
.bin455 MB · 63%
From the Hugging Face model README
De-noising Diffusion Implicit Model trained on teticio/audio-diffusion-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference.