StemRoller – Isolate vocals, drums, bass, and other stems from any song
github.com
github.com
If I didn't know who did the real work and benefitted a lot from this tool, I'd give to StemRoller in proportion to my gratitude -- which I'm sure others are liable to do.
How about saying so above the fold?
Saying so and suggesting an alternative isn't being pedantic.
It does seem rather disingenuous that your replies make no mention of admission for making a factually incorrect statement in the first reply.
Not a big deal to me personally, but it is not surprising that some people see this as being petty.
To the downvoters above: is a "thanks" to a correcting comment not enough on HN?
Also, was it rude to say hiding content "below the fold" is a design problem?
This is a very odd thread to me - like I was being chased down by pedants who are calling me pedantic.
I’m definitely happy to see more front ends for Demucs being developed and to read that it has been useful to other musicians!
We are working on the next iteration of the model, and with more sources, hopefully released by the end of the year :)
If you are interested in this research you can follow my Twitter (@honualx) or star the Demucs repo.
Is there any chance you can disentangle guitar and keyboard? I work a lot with Grateful Dead music and I'd like to be able to pull jerry's guitar out from the keyboard from live shows. Similarly, it would be cool if you could parse shpongle into its consituent tracks, but I think that's probably impossible.
The backing vocals seem to have disappeared for the most part, and are only audible in the vocals stem when the lead vocal is present (like they're reverse-ducked? Been a while since I did any production, the terms have escaped me...).
Not much use with complex arrangements to be honest, I was hoping to get things like the strings section separated from the rest of the arrangement.
Original: https://www.youtube.com/watch?v=tWX3El-slpY
Output: https://file.io/etpOQt57ziKe
(Only asking because you linked to YouTube, and I'm not sure if you used the YouTube audio for your source.)
I typed the song in the search and pressed the first likely result, which is the youtube video I linked. Using the software as intended I believe.
I split https://www.youtube.com/watch?v=DDaL7KBjkDI
And it gave me this https://www.dropbox.com/sh/inyk38n2jrp5i45/AACpB0xXNFxamEmP3... I noticed some weird hissing with the 808s, but other then that it sounded pretty good
For more of a challenge, I inputted https://www.youtube.com/watch?v=uAwQ3njiU4M
and it came up with https://www.dropbox.com/sh/97lzke0puh9dzeo/AACE75vsbNS43UqqH... It was able to separate some of the kicks from the 808s, which is really impressive to me!
Overall, I'm very impressed! This sounds much better then lalal.ai to me
You also have to build the demucs-cxfreeze dependency (as described in its repo, https://github.com/stemrollerapp/demucs-cxfreeze).
Trying it out with Alan Walker's Alone, it separates the vocals and drums almost perfectly. Bass is really fine as well, only instrumental and 'other' was a bit mixed up in my try.
Why? Why can't this just point to the location where ffmpeg is rather than making a copy of ffmpeg? symlink might work, but just do a $(which ffmpeg) or ask the user for the path ~/bin/ffmpeg /usr/local/bin/ffmpeg etc
[0] https://www.openculture.com/2022/04/hear-the-beatles-abbey-r...
Microphone bleed, lots of overdubs (especially vocals), and repeated re-layering tracks on tape over and over due to channel limitations. They really were doing crazy stuff with limited tech.
I think this would be hard for bands that really fill the spectrum and don’t have that clean treble, mid, bass separation. Or recordings really compressed into a frequency range.
Now this makes me want to see what happens with like My Bloody Valentine and Husker Du :).
Wait fifteen minutes and out pops four stems, flawless so far, even been messing around with mainstream tracks and using ableton with warp applied to quickly build out remixes. Demucs is going to be /is already a game changer!
1: On second thought maybe not. It has not aged well.
2: Me and another kid, with a guitar, a pre-OS X Mac, a pirated copy of Rebirth, a pirated copy of SoundEdit 16, and literally the mic that Apple used to include with (some?) Macs. I’d back-reference[1], but our equipment was not the problem. Well, except for [3].
3: I learned my lesson: I should have been older and had a job that would afford me a backup drive, so I could sample the sounds of that dying HDD and retcon the samples into my “album”[1].
Seems like this tool might be better than Izotope's.
in a semi blind comparison, I prefer demucs for all 4 tracks (drum, bass, vocals, and other). bass and other stand out the most so let me say a couple words about them.
bass - the demucs bass has less bleed from other instruments and the volume is consistent throughout. with spleeter, the volume varies a lot and there are multiple sections of 1-2 bars where it just drops out completely. In Capo, the demucs spectrogram is nice and clear whereas spleeter tends to look like pencil smudges for the most part.
other - with spleeter, whenever there are vocals, the other instruments turn to mush. demucs is much better. Oh, you can tell people are singing -- the instruments get muffled -- but you can still hear them.
I like the opportunity to view the source code and learn from it, as opposed to most paid products which are typically closed-source and a bit of a "black box".
Sure - this is mostly just an accessible frontend for Demucs, but that's still okay. The author clearly indicates that in his repo, giving credit where credit is due. Additionally, this helps less-technical creators be creative in new ways.
Thanks to all who contributed.
If the elements of the song are recording in isolation - which they are in all studio versions, why can't we just move to a format that supports the layering?
Say that you want to make a remix, mashup, or otherwise use sound-bytes from a song. The easiest thing to do is use a tool like Spleeter/Demucs to separate the source layers so that you can then further process them in your DAW.
This is what I do, but I just use the Demucs CLI because it's simple enough.
As after all the sound quality doesn't interest me too much to do this, I usually use iZotope RX, but I will try this tool.
Yes, I agree.
Else who cares