Downloads · 30 days
10
0% of all-time downloads
teticio/audio-diffusion-breaks-256
audio-diffusion-breaks-256 is a machine learning model from teticio. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for diffusers.
Denoising Diffusion Probabilistic Model trained on teticio/audio-diffusion-breaks-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The audio consists of 30,000 samples that have been us…
Downloads · 30 days
10
0% of all-time downloads
All-time downloads
2.4K
Public
Repo size
6.5 GB
Likes
4
Public
Click a slice to open those files.
.bin455 MB · 62%
From the Hugging Face model README
Denoising Diffusion Probabilistic Model trained on teticio/audio-diffusion-breaks-256 to generate mel spectrograms of 256x256 corresponding to 5 seconds of audio. The audio consists of 30,000 samples that have been used in music, sourced from WhoSampled and YouTube. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference.