Downloads · 30 days
53
0% of all-time downloads
teticio/audio-diffusion-256
audio-diffusion-256 is a machine learning model from teticio. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for diffusers.
De-noising Diffusion Probabilistic Model trained on teticio/audio-diffusion-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The code to convert from audio to spectrogram and vice versa…
Downloads · 30 days
53
0% of all-time downloads
All-time downloads
15.9K
Public
Repo size
5.7 GB
Likes
7
Public
Click a slice to open those files.
.bin455 MB · 69%
From the Hugging Face model README
De-noising Diffusion Probabilistic Model trained on teticio/audio-diffusion-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference.