The music project that I have been working on aims to bring together a musical audio pre-trained language model, natural language inputs, and a novel interface that deconstructs what you hear into visualizations and user controls. The idea is to allow you to dramatically improve personalization of music search, discovery, and curation by allowing interactivity with the individual parts of music, also known as stems in music theory and music mastering.
The overarching challenge that music streaming services and more broadly media streaming services have, is building a sustainable business on somebody else's art...you have to (rightly) pay licensing costs. This is a major (and ever changing) business risk and one of the larger challenges when I think about launching a new music streaming platform. MTV discovered this in the mid 80's when margins for music and music video licensing became flat, Netflix / Apple / Amazon / etc know this which is why they invested in developing their own content in house.
If anybody is passionate about working on something new in streaming music let me know - I could really use help developing the music audio PLM prototype.
(1) https://www.musicbusinessworldwide.com/there-are-now-120000-...