Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
https://github.com/lucidrains/musiclm-pytorch/blob/main/musi...
https://github.com/lucidrains/musiclm-pytorch/blob/main/musi...
i assume there's only a superficial description of the architecture, and no weights to load in, so you'll have to train everything from scratch? do we even have their dataset?
[1]: https://github.com/lucidrains/denoising-diffusion-pytorch