Downloads · 30 days
0
marcop/musika_ae
musika_ae is a machine learning model from marcop. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for keras. The card lists the license as mit.
Pretrained universal autoencoder model for the Musika system for fast infinite waveform music generation. Introduced in this paper.
Downloads · 30 days
0
Access
Public
Updated Oct 24, 2022
Repo size
113 MB
Likes
5
Public
Click a slice to open those files.
.h5113 MB · 100%
From the Hugging Face model README
Pretrained universal autoencoder model for the Musika system for fast infinite waveform music generation. Introduced in this paper.
The Musika autoencoder consists of two hierarchical stages that are separately trained. This autoencoder is trained to encode and reconstruct general 44.1 kHz waveform music. The final time compression ratio that is achieved is 4096x. As an example, 23 seconds of 44.1 kHz audio are encoded into a sequence of 256 vectors with a dimension of 64.
This autoencoder is automatically downloaded and used at the first execution of the system. Try Musika here!
The autoencoder was trained on both the SXSW dataset (diverse music dataset) and on the VCTK dataset (speech dataset) to produce general representations.