CoolFace
Datasetpublic

teticio/audio-diffusion-1024

Over 20,000 256x256 mel spectrograms of 5 second samples of music from my Spotify liked playlist. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference using De-noising Diffusion Probabilistic Models. x_res = 1024 y_res = 1024 sample_rate = 44100 n_fft = 2048 hop_length = 512

sourceHugging Faceupdated 4y agoView on Hugging Face
0likes299downloads
Dataset Card

Over 20,000 256x256 mel spectrograms of 5 second samples of music from my Spotify liked playlist. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference using De-noising Diffusion Probabilistic Models.

x_res = 1024
y_res = 1024
sample_rate = 44100
n_fft = 2048
hop_length = 512