Apple Music Sing
apple.com
apple.com
> Apple today announced Apple Music Sing, an exciting new feature that allows users to sing along to their favorite songs with adjustable vocals1 and real-time lyrics.
Users can already "sing along to their favorite songs" -- this feels like a wasted opportunity in the first sentence to explain the differentiator. It feels like they're misusing/misunderstanding what "sing along" means, and the real feature enhancement here is that users now have the ability to suppress vocals in songs. (Which is kinda cool.)
I get that this is a press release, but "show don't tell" would go a long way in announcements like this, even (especially?) if it's not available yet.
Another thing is that when you sing a foreign song, the karaoke also means to read the foreign subtitles in your own language. It's an extra meaning they might want to avoid, I assume this is not a feature they have.
It also might imply perfect vocal removal, which they aren’t guaranteeing.
"Hey we just added a cultural pastime as a feature!"
There are some that could twist this feature into "Apple profiting off of cultural appropriation".
1) Perhaps they don’t want to trigger a different licensing model (on either the music or the lyrics) that applies specifically to karaoke devices and venues. Just a guess, not a music lawyer, but all sorts of unexpected terms apply to music licensing (such as streaming TV services not being able to use music from the original TV broadcast) 2) Perhaps they don’t want to associate with the karaoke “brand” which some users might perceive as kitschy or low-quality (have you ever seen a karaoke video?)
And "real-time lyrics" is _already_ in Apple Music.
In your/Apple's defense, the photo they included is pretty clear what's going on. Important to note the footnote on the page -- this isn't even full-fledged karaoke: "The vocal slider adjusts vocal volume, but does not fully remove vocals."
I'm being nitpicky for sure, but that's kind of what HN and similar forums are for. Especially when it's Apple, which I hold to a pretty high standard for product marketing.
And "real-time lyrics" is _already_ in Apple Music.
Currently, Apple Music lyrics are synced to the current line of lyrics (if the lyrics metadata supports it)This new feature appears to sync the lyrics to the current word of lyrics, which is what karaoke usually does. Not sure how big a deal that is to folks looking to do their own vocals.
I'm curious about how they're doing that word-for-word timing. Hardcoded metadata? ML? Is it a simple interpolation based on the existing line-by-line timing?
I'd guess it's done using the same speech-to-text system used by voice assistants, which can certainly show the words it hears in near-realtime -- and way more quickly when it already knows what words it's listening for.
By the way, karaoke often highlights individual syllables, not just whole words.
edit: ultimately I could care less (but not by much) about using it; I am just curious about what they offer and how they compete.
I know the last thing I want is for some random person to sing the song we're all listening to given that most people can't sing.
I find myself usually telling the kids in the car to pipe down and let the artist sing.
So I'm really interested in understanding the demo for this feature.
It seems that your kids are the intended target.
Kids love to sing along with music. You've already discovered that.
You must not have 10 to 12-year-old girls. Put a couple of them in a room, and eventually they'll start singing.
Wondered if "karaoke" is trademarked? (or patented?!)
It doesn't seem to me likely that they didn't think their audience would know the word "karaoke" at this point, but I don't know.
And if Apple knows anything is how to market / rebrand the same old shit and sell it to the masses at really marked up prices for profit.
Polish and perfection makes this less appealing to me. Makes it feel up-tight and a bit snobby.
Also, I had a friend who got addicted to Smule which is an online sing duets with random strangers. It uses auto-tune so that anyone can enjoy themselves. I never tried it but he said it was amazing and he got a ton of joy from it.
Me, I like the idea that I can show off my amazing singing skills so if this auto-tunes everyone then that will go away. On the other hand, it's always been frustrating that box style karaoke places rarely have a large selection of western music (most of them are targeting the subset of their local market that would actually to got a box style karaoke place). My hope is that with this more of them will support just using the room as a place to use this service and then my friends and I can sing from a larger variety of songs.
I don't know if this helps at all, but you probably want to learn to sing in a different octave, not a different key. This means shifting your voice up and down from the original in multiples of eight notes. You can sing in a different octave without changing the background music and it'll sound completely normal, and have the same feel as the original song.
https://www.apple.com/jp/newsroom/2022/12/apple-introduces-a...
They're also pretty good at taking a product that is about 75% of the way to being great, and pushing it that last bit. Hell, that's arguably why they are so successful. They're not really inventors, they are perfectors.
One notable difference is that with Apple Music Sing you're singing to the original track instead of an instrumental reproduction.
Everyone I know would call that karaoke. And I assume will when telling their friends about this new feature.
If I search for Apple Sing on google news, the first 5 articles all call it "karaoke-like" or just "a karaoke feature".
Whatever reason they have for not saying the word karaoke, I don't think it's that it's not what most people consider karaoke.
A microphone (and ideally having your own vocals mixed with the audio output) is the core feature of karaoke, not playing an instrumental version of the song.
But this has become a pretty silly "debate".
Me, I'm still curious why Apple did not use the word karaoke, and do not myself think it's becuase they don't think people will consider it karaoke, but I understand you do, cool.
I think most US courts would see "karaoke" as a generic word today, but I don't know about courts outside of the US, and Karaoke is still seen as a part of the "trade dress" of a couple specific US corporations even if they can't entirely protect that trademark, they are still allowed to fight for it. At least a couple seem to still exist primarily as trolls for lawsuits today.
According to the toy aisle of US stores, "Sing" (and "Sing Along") is the only real generic English term in use, and very few (no?) current toys use the word "Karaoke", so I'm guessing those trolls are still winning in the US, for now. (Again, I don't know about the rest of the world, and toy branding is an anecdotal source at best, but still an interesting proxy view in how much large companies assume they will be sued.)
It's what the marketing machine at Apple does.
It was explained to me that the reason YouTube didn't offer a karaoke feature -- despite having licenses to a lot of lyrics -- is that karaoke is considered a separate license.
Even if you license both the recording and the lyrics, combining them into a karaoke feature isn't on the table by default.
I personally have no idea how accurate that is, but this was the scoop I got from those in a position of authority.
If it is indeed true, then perhaps that is a factor in Apple's decision not to use that word.
For prior art, see what happened to Aereo. A little-known fact is that YouTube TV started with the exact same strategy. But Google, of course, had more money and lawyers to get over the hump.
Given the monumental size of fastidiously annotated music corpuses available, the quite diligent and far more agreed upon systematisation of various forms of scientific description of musical forms, and the arguably simpler space than some of the more recent advances like 3d object generation and the sophisticated artistic output from the latest 2d image generating models... seems rather odd music has no wildly popular machine learning model, something i could ask "give me 5 minutes of synth-wave by Mozart" or "Prodigy's Firestarter, but without the lyrics"...
... My suspicion is that the entire endeavour is tortured by licensing and at every turn must avoid ever sounding like music that could be owned by someone else, and as a consequence cannot be "popular" as it must be crippled and limited, built to at all costs avoid ever sounding like Taylor Swift, Queen, Aerosmith, et al.
edit: I missed the "adjustable vocals" in the list of features. No idea then. Footnote says "The vocal slider adjusts vocal volume, but does not fully remove vocals."
This Apple feature seems to be where the same popular recording gets used, which has vocals, and then some magic processing reduces the vocals.
There has always been software to sorta-kinda suppress vocals from existing recordings, but it usually ends up sounding like trash (artifacts of all sorts) compared to just having a karaoke band re-perform the song (although the talent and production quality can be dubious with the latter). Presumably the innovation here is not having vocal suppression sound so bad.
Aside: and then there are popular artists who release an instrumental version, or stems from which an instrumental version can be trivially created, but this is so extremely rare as to not even be on-topic for a conversation about broad availability.
To me karaoke is just signing to music, which is exactly what this is. I found it very strange that they avoided that word.
Properly licensed karaoke tracks are covers, because licensing custom remixes of original tracks is prohibitive. There are sites where you can buy stems/backing tracks of any song you can think of, which are reproductions by talented musicians.
https://www.karaoke-version.com/custombackingtrack/wham/last...
Regardless of whether you believe it's an "essential property", that karaoke music often sounds cheesy is one of several reasons Apple wouldn't want Apple Music Sing to be associated with it.
Apple's footnote disclaimer says "does not fully remove vocals" and "kara" means "empty/void" so it kind of makes sense to avoid using such an absolute term to describe something that isn't.
No, this just isn't what the word means, either popularly or in technical usage. I can't find any dictionary that restricts karaoke to covers, and if anything the dictionary emphasizes non-covers:
> an act of singing along to a music video, especially one from which the original vocals have been electronically eliminated.
https://www.dictionary.com/browse/karaoke?r=75&src=ref&ch=di...
It's true that in practice most karaoke audio tracks are covers for financial reasons, but that doesn't change what the word means. Most tables have four legs, but six-legged tables are still tables.
I agree that it's not what the word inherently means, but it's the typical situation (which is effectively like meaning) where I am (USA).
See the many examples (in the West) of tools that remove vocals from audio tracks; they are ubiquitously described as useful for creating karaoke audio tracks:
https://tunebat.com/Vocal-Remover
that's more the definition of "cover" than "karaoke" (which is a live sing-along at a bar or at home)
The real surprise is that the artists are ok with this, they must be getting compensated.
Then it would seem even more like "Apple Karaoke."
I'm thrilled to see this announcement and will likely buy my first Apple TV because of it and switch my Pandora subscription to Apple Music.
As long as you have a 0 latency Mic setup then just that and YouTube searching Karaoke versions with AdBlocker works beautifully.
At the end of the day, I just didn’t want come right out and say “rsync, ftp, something something Dropbox”. :-)
Routing the mic through your AV setup _will_ result in a delay which’ll make for a horrible experience. So don’t bother. Get a cheap mixer and a decentish Shure mic, a single active studio monitor and you’re away.
A smart tv, a microphone and a speaker? Or get a karaoke microphone with a built in speaker.
https://knowyourmeme.com/memes/i-made-this
While I have zero interest in karaoke, I imagine it's devastating for a bunch of karaoke apps that are now going to see new sales evaporate. I don't think it's cool for the company to just take over a market, or a smart thing to do when their app store monopoly is a political football. I guess they're relying on antitrust enforcement being more of an idea than a practical reality these days.
- Young artists can (illegally) use beats by popular producers. Young producers can listen to various tracks and better understand some techniques that were shadowed by vocals before.
- More bootleg remixes of popular tracks. Things that weren't possible by just slicing and sampling the track would suddenly be possible, too.
- Better music stemming / demixing ML models. I believe right now the baseline is Spleeter by Deezer [1], which is pretty good but leaves so much room for improvement.
Tons of reasons it sucked: stereo reverb on vocals wouldn't cancel out, other instruments panned dead center (or at least highly correlated) would cancel out along with the vocal, and early MP3 encoders often sounded better in "joint stereo" or "intensity stereo" mode (assuming low/mid bitrates) which very reasonably gave more bits to the center and fewer bits to the sides, so the differential sounded like hot garbage.
Exciting to see that this technique has been replaced with ML.
Of course people can still rip it off an iPhone/iPad through the "analog hole" via the Lightning to 3.5mm adapter
It works pretty well on Mac
It does work with music you have purchased on iTunes.
So it's an IP issue. It also works with the higher tiers of TIDAL or SoundCloud - https://help.algoriddim.com/hc/en-us/articles/360012493060-W...
The reason I say that is because I suspect they are doing this with a neural net.
Splitting tracks on-device would probably use a neural net, but I don't think they're doing that.
Then in the mid-2000s there were a few attempts to use algorithms for this (that mostly failed) and circa-2015 we started to see machine-learning based options like Spleeter and Demucs. This is what Apple is implementing here, and it's been freely available for years. It's not really novel for anyone who isn't writing their music on an iPhone.
I mostly ask because I've spent the past 2 or 3 weeks working on a remix album, and haven't really had trouble finding stem separation tools that work well.
https://www.loudersound.com/news/metallica-and-justice-for-a...
Can anyone make out if songs like this are specially mastered with separate vocal channels? I’m guessing they’re not minus ones that someone else has arranged and played. If these are coming from the original artist it’ll be interesting to know if they’re being uploaded as separate vocal channels or if Apple is applying ML to do voice isolation.
Do any audio file formats support separate tracks with level info?
> Do any audio file formats support separate tracks with level info?
I don't know what file format audio engineering software uses to store all the track info but my guess is that it's akin to a zip file containing a flac for each track.
I've never had this produce good results. It removes major parts of the music, while also not fully removing the vocals.
I would actually take that and record out from that to another tape, so I could have 'instrumental' versions of the songs.
And, no - it doesn't work very well at all; it sounds like a poor 64kbps MP3 with bizarre artifacts and; yes - many missing instruments.
Presumably if they have Don’t Stop Believing and a separate karaoke karaoke today, they could basically cross fade between them for the adjustment of original singer vocals. And it would work better than trying to remove vocals after the fact.
But that would require both versions, synchronizing them, and KNOWING which two tracks went together.
Is that easier or harder than getting the labels to just give you a special version with an extra channel for vocals?
Hopefully someone digs in and finds out once this is released.
Depends on the label and the artist. Could be easier, could be much harder. Sometimes multi-stems just don't exist anymore; were lost, have additional licensing issues; labels are a nightmare to deal with.
Almost assuredly this is just ML-assisted frequency separation.
It absolutely could be done, I just think that Apple would want very good coverage and studios would be very slow to provide this format.
It's already become pretty common for studios and labels to make stems available (stems are full multi-track files that can be used with a DAW) to industry insiders and even the public sometimes. There's a community of remixers, samplers, even a small cottage industry of YouTubers who work with these regularly. It wouldn't take them more than a few minutes per track to annotate which channel is vocals.
Do you stop at instruments and vocals or are we talking 100+ tracks? It strikes me as nontrivial work for very limited purpose. Apple can turn your vocals down so now your exports have to include extra tracks? I don't think that's a good enough sell. The labels do care about this use case.
It makes less sense when the intent is to mix those tracks together and send them to a single output device. Producers and audio engineers want full control over that mixing process because it's almost never just a simple sum. They are doing audio compression (sometimes multi-band), dynamic EQ, saturation, limiting, etc. They wouldn't want to give that up, because it's an essential part of making a good sounding recording.
"Coupled with an ever-expanding catalog that features tens of millions of the world’s most singable songs"
"The vocal slider adjusts vocal volume, but does not fully remove vocals"
Based on the quantity and results it's ML like Spleeter (https://research.deezer.com/projects/spleeter.html).
There is a 4-track an audio format developed by Native Instruments that is mostly used by DJs, and only dance music gets released in that. However, now that I search their website, there's no mention of it anymore — guess it wasn't popular enough with the labels to get any traction.
Though I have no idea how to get into this beta test.
This is a killer feature for me as I am trying to learn to sing.
There are some services online that use AI to split tracks into instruments and vocals that do hell of a job, but trying to combine them with written lyrics again is a pretty painful experience.
> The feature won’t see users switching over to music tracks that already have the vocals removed, however. Instead, it’s relying on an on-device machine learning algorithm that processes the music in real time, Apple says. The algorithm isolates the vocals from the rest of the song, allowing users to adjust their volume accordingly using a new slider button in the Apple Music app.
(TechCrunch)
I wonder how well this will work. Especially when some songs apply all sorts of wacky effects on vocals. Sometimes it’s not even clear to a human what should be considered vocals and what’s part of the instrumental. The context matters too. Maybe they’re using the lyrics as a signal?
Having easy access to instrumental versions of millions of tracks will be huge for remixers, people recording vocal demos, etc.
...well, maybe.
I'm sure you won't be able to export the vocal-less versions out to an audio file directly.
But you should be able to turn off the vocals and capture the resulting audio by the usual means: loopback audio drivers, analog capture, etc. It's still a step, but it's going to be fairly easy and the resulting file should be quite clean.
(It's also going to revolutionize the singing I do alone after a few glasses of wine but that's not super impactful to anybody but me)
The easy way would be if it knew which tracks were instrumental or not and only play those. The GREAT way would be if it just played everything but knew how to remove the all vocals by playing just the backing track.
Is that still the case?
I am assuming, perhaps incorrectly, that Apple has access to multitrack versions of some songs and that for these tracks it will be able to remove the vocals in a pristine way.
I wonder if this will have the real killer feature of professional karaoke ... the ability to transpose the song into a more convenient key for the singer. Apple have had this capability in Logic Pro's pitch shift and it's pretty good even on polyphonic music. Some Karaoke platforms use midi and midi-like formats and synthesize the instruments to do it, but I'd be surprised to see that here.
Wasn't what I was looking for initially but it was surprisingly good compared to the older "vocal remover" type plugins I remember having for Winamp et al back in the day. If I ever actually DJed anymore it would definitely be useful for mixing tracks.
Nope - it's using ML.
Source: https://techcrunch.com/2022/12/06/apple-music-is-getting-a-n...
I switched from Spotify to Apple Music about a month ago, mostly because I wanted music in Dolby Atmos. (Not all songs are in Dolby Atmos, but many are, and they sound far superior on my AirPods to anything without it.) After transferring my Spotify playlists to Apple, I've found that the experiences are pretty similar. There are a few things I like better about Spotify (e.g., it has better support for non-Apple platforms like Linux and Windows), but I prefer Apple now. If you're curious, I'd recommend a free trial.
1. The Apple Music desktop app is utter garbage. Other than performance issues and crashing, which happened to me almost daily, the UX is a decade behind. It's dead simple things... like the artist's name on the currently playing track isn't even a link.
2. Spotify's recommendation engine is outrageously better than Apple's. With Spotify I find new music that I actually like almost weekly, with zero effort at all. All Apple does is pigeonhole you with whatever you've listened to recently, I had a day where I listened to "lo-fi chill" music and that's all it recommended to me for weeks.
But the old Apple TV 4K is not powerful enough for displaying text on a TV, even though any old iPhone is?
I get that they want to push device and service adoption but they should pick a lane, either push users to upgrade devices, or push to sign up to paid services. Pushing for both at the same time is greedy.
Using GPU's for ML is not very efficient on low-power devices. Having dedicated ML circuitry is the way to go.
Edit: per the bottom of the page...
> Apple Music Sing will be available on all compatible iPhone and iPad models as well as the new Apple TV 4K.
:/ maybe they do ML to process stuff on-device, per other comments' speculation, that justifies requiring the newer model?
In addition, the Apple TV is plugged in, it doesn't have to limit its power for battery reasons.
This seems like it might be Apple‘s standard thing of just promoting whatever their newest products are in all communications.
...which leads to users upgrading their devices
I mean, it's not a bad marketing strategy when most people who can afford an apple device already have 1+
So I wouldn't be surprised if it relies on a particular hardware chip that the older Apple TV simply doesn't have. That has definitely been the case for everything Apple has launched with regards to Spatial Audio.
Just curious: what in the article makes you think that?
Look at the list of what Apple Sing includes:
Adjustable vocals: Users now have control over a song’s vocal levels. They can sing with the original artist vocals, take the lead, or mix it up on millions of songs in the Apple Music catalog.
Real-time lyrics: Users can sing along to their favorite songs with animated lyrics that dance to the rhythm of the vocals.
Background vocals: Vocal lines sung simultaneously can animate independently from the main vocals to make it easier for users to follow.
Duet view: Multiple vocalists show on opposite sides of the screen to make duets or multi-singer tracks easy to sing along to.
------
The part of the article where they state these things explicitly...
As a matter of fact, calling out "millions of songs in the Apple Music catalog" actually makes it seem like the adjustable vocals will only be available on certain songs that they've added support for.
My guess is that it's entirely dynamic. It's hard to imagine the complexity of doing batch processing to render each song in the library, and maintain that as new songs are uploaded, and update the renders for software improvements. Better to just do it realtime.
And since classical, instrumentals, esoteric ambient stuff, death metal, etc, will probably not be supported by the algorithms, I think the "millions" refers to those that can be processed in realtime.
> [...] an on-device machine learning algorithm that can process music in real time
Their latest apple TV includes the A15 which includes a 'neural engine' for ML, and this is also included in their latest iPhone / iPad, so that might be part of it.
I think this only requires pre-making two audio files per track, and simultaneously streaming these.
Real-time lyrics, Background vocals and Duet view are all nice features too, but the hardest part processing-wise is analysing how loud you sing into the microphone. It's just karaoke with a good UI.
They are likely using a sophisticated ML version of what old karaoke machines did, and removing the vocals in real-time.
Source: https://techcrunch.com/2022/12/06/apple-music-is-getting-a-n...
Wonder why they take this approach though, as it is clearly over-engineering (if I correctly understand that the goal is just to make vocals volume adjustable).
Depends what the other non-functional requirements were. i.e. if the NFRs were as follows:
* Cannot increase bandwidth / mobile data usage.
* Cannot impact music quality / bitrate.
* Has to work offline.
* Cannot increase on-device storage.
* Has to be responsive.
Then two audio streams might not work.
Another advantage of doing it on-device is that it doesn't actually change any of the backend architecture too. It might be a lot of change to a lot of systems for a feature which only adds a small amount of functionality - i.e. architecting your entire backend and streaming around seperating audio tracks might not be the right focus.
That’s the understatement of the century. “this only requires […] simultaneously streaming these”
Since Apple is all about on-device processing with so many of its features, going back-and-forth to the data center doesn't seem to be its style these days.
That's more of a Google thing.
And no one can accuse Apple of telling its advertisers that you start your day with Funky Cold Medina.
It makes it so you have to buy new devices sooner.
The privacy thing is a nice side-benefit and PR thing, but let's be realistic here.
EDIT: Just to remind everyone, we are literally in a thread about a new device feature that is trivial to do in the cloud, which Apple chooses to do on-device, which makes it only a feature for its newest generation of products...
Is that the actual reason though? My personal impression has been that it's a combination of reasons that benefit Apple. The increased user privacy being a nice bonus for users, but not the primary reason:
1. Producing phones powerful enough for on-device ML both justifies the high price point to the general public and is a good marketing point (along with increased user privacy)
2. Avoid backend infrastructure costs. Why spend extra money on servers, maintenance, and compliance when they can just offload the work to the devices themselves since they're capable?
3. Bonus: The unplanned obsolescence for new features like the one announced is also a side effect that benefits Apple.
I do not get the impression that Apple's primary focus is to benefit users and their privacy.
If Apple wanted to support this for user-provided mp3s then on-device would make sense. It doesn't sound like they support that though.
It requires new screens and display controllers that can refresh at much lower rates, and stronger guarantees of burn in prevention.
Like sure, you could do it on an older screen but you’d burn through battery much quicker and potentially damage the device in the process.
This is true for Android devices as well.
I am getting fed up by the user hostile attitude of producers and providers, maybe I will just keep my money or spend elsewhere.
Weird kind of devotees.
Pushing users to upgrade devices is unethical and harmful. Apple does it all the time and they should be held accountable for that.
Discarding old devices and creating new ones destroys the environment and increases pollution. The best device is the device you already own!
Their successful business strategy results in environmental destruction and other issues. Wh
But ignoring that for a moment, are you proposing that companies shouldn't improve hardware because it leads to software features that don't work on old devices, which leads to old devices being thrown away? That doesn't feel right to me.
[0] https://www.sciencedaily.com/releases/2018/10/181016142434.h...
During the lockdown, I had several Apple products that I wanted to bring to the store for disposal, but Apple wasn't taking any devices at all.
Then, once the restrictions were lifted, I found out that only one of the six Apple Stores in my area takes devices, and it was an hour away.
It seems to me that every Apple retail location should take back devices.
Huh. Long-time Apple user and my experience is they're really good about supporting older hardware, including when that hardware can't support the newest features. Usually (not always, but usually) if an older device doesn't get a feature, it's because it can't, but they keep getting updates anyway, just without the features they cannot support. That's the trade-off for more features being local-only or local-mostly rather than "cloud".
Honestly, it might not be. My Apple TV 4k seems slow as molasses with performance more like an AMLogic 905.
Not sure if it’s because it doesn’t have enough RAM, the flash is slow or the CPU is just way under powered. Yes, they advertise it as an A12Z or whatever, but it’s a binned part that wasn’t acceptable for a phone.
I currently own all three generations of the Apple TV 4K. All of them are snappy devices.
Now that said, I use them for streaming video only, both a variety of services as well as use Infuse for streaming from my NAS.
As someone that works on set top box / streaming chips, I don’t think you realize how low that bar is :P
The prev gen ATV 4k is clearly superior, I’m not disputing that. But ‘snappy’ is a relative term. I expect a lot more out of a device sporting the chip it has. I regularly get input lag, stalled apps (that the system cannot recover from for some time), bursted input (like it doesn’t debounce queued remote inputs after stalling)
There are systems level issues that remain unaddressed, and I speculate they are due to hardware bottlenecks somewhere .
The note at the bottom reads: "Apple Music Sing will be available on all compatible iPhone and iPad models as well as the new Apple TV 4K".
Maybe it requires the Neural Engine that was added with the A11 and opened to third-party apps with the A12. (Sep 2018).
But that should still allow the Apple TV 4K (second generation) which uses the A12.
iPhones/iPads with the A12 or newer are the iPhone XS/XR, iPad Mini (5th generation), iPad Air (3rd generation), and iPad (8th generation). The iPhone SE (2nd generation) uses the A13 so that should be compatible too.
If you use the Help menu's search feature it will direct you to the “Show Lyrics” option, too.
At first, I thought it meant that this only was on the current Apple TV 4K, but then I thought it might just be marketing speak to get people to buy Apple TVs.
But maybe the jump from the A12 to the A15 does provide something that makes this possible. The only thing I can think of is the neural engine (although that first debuted in the A11).
What is cool about HN though is that maybe an Apple engineer will show up anonymously and give us the low down about why. In my short experience at Apple there were plenty of problems but NOT shipping things to customers to make money was not even remotely one of them. Bending over backwards on ancient radars to support the iPhone 6 sucks up an enormous amount of time!
This Music Sing service might also be really good at parties.
There are lots and lots of existing providers of karaoke tracks: these invest (depending on quality) between several dollars and several hundreds of dollars to record "soundalike" tracks
Sales of such tracks do not generate royalties for the original performer, but do pay out to the composer (per track sale and for things like public performance).
Apple is now garroting these middlemen using technology, and most likely using this capability as leverage in negotiating with recording artists ("hey, give us a 14-day exclusive on iMusicOrWhateverWeCallItThisWeek, and we'll kick back an additional point on residuals").
This is bad news for the existing providers, and barely good news for anyone else.
Who Moved My Cheese comes to mind.
There are lots and lots of existing
providers of karaoke tracks:
these invest (depending on quality)
between several dollars and several
hundreds of dollars to record "soundalike"
tracks
Commercial karaoke establishments will still need to pay for the licensed versions if they want to be legal. That doesn't change.People who don't want to do that already had tons of options - Adobe Audition and tons of other software can remove/reduce vocals.
So I don't feel like this changes the commercial picture too much? I feel like this will mainly affect at-home singalongs.
The at-home market in the UK, on the other hand, is pretty significant. None of these households know how to operate Adobe Audition or something similar: they just want to sing along with whatever is on the telly*.
There are lots of companies catering to that market. In fact: in the past, Apple was more than happy to allow them on their platform, to fill in the gaps left by Apple's inability to negotiate certain agreements.
In the past year or so, Apple has gotten more and more restrictive with regards to "soundalike" content. And we now know why... Is this inevitable? Possibly. Is it fair? Maybe. Is it yet another cottage industry that Apple strangulates? Definitely.
*And yes, this is a very simplistic caricature by choice. Of course UK consumers are more sophisticated, but...
I didn't realize there was a commercial home market for those professionally produced karaoke tracks.
If I'm understanding things somewhat correctly now, this does sound like it will be a large blow to that market.
In what cases to we continue to use a slow, antiquated process to do a task when an automated technology comes out that can do it faster and/or better?
Sure, I could pay a contractor to go through my files to find and replace every instance of a text string with another text string. Or I could use sed. It's not anti-competitive for sed to exist.
I don't care if their audio processing to drop the volume of the lyrics isn't perfect. It will give my daughter the whole Apple Music catalog to sing along with and without the side of price-gouging.
For example, anyone remember those Disney sing-a-longs with the bouncy mickey mouse icon? IMO, the icon really helps communicate timing information; you can tell what's coming by the speed and arc of the icon. I'd love to have that system with a wider variety of music.
There's also RockBand's approach, which displays scrolling syllables to be sung when they cross a white line. It's a bit harder to read, but I'd take it over traditional karaoke.
After all, it makes sense. If you want to engage subscribers to your service, what else would you do? How most people engage with recorded music? Well, they just… sing along…
I'm also curious (like some on this thread) how they are technically separating the vocals. Is this yet another proprietary music format that artists have to mix/master to? (Like spatial audio)
I also really miss the days when Singstar was a thing, or even UltraStar, the open source “clone”.
So when can we get an Apple Arcade Rock Band/Guitar Hero type game that includes vocals, guitar, bass, keyboard and drums and works with any song on Apple Music?
And how about an automatic Dance Dance Revolution/Just Dance/etc. with the entire Apple Music catalog?
The downside during voting season is that Disney songs are interrupted by rape allegations that 6 year olds ask about.
I'd love to get better at singing, but not quite enough to actually find, schedule and pay for singing lessons from a human teacher.
But I'd happily do 15 minutes a day with a Duolingo-style app, especially if it could guide me through exercises to help me sing better.
I'd equate being able to speak a foreign language to singing in key. That's what the apps will be able to help with, at least from what I've seen. If that's what you're looking for, that's a good place to start. However, there is a lot of technique to make your voice sound pleasant and interesting beyond just singing on key. I'm not sure if that apps really help out there that much like an actual vocal coach could.
https://www.microsoft.com/en-us/research/project/mysong-auto...
Not perfect but it's pretty damn good.
Our Amazon Echo Show started doing this recently when listening to songs on Amazon's own Music service. I noticed it, thought it was interesting, commented once to the wife, then went back to just listening.
As a user/listener I don't really care. But the engineer in me is curious.
To think its coming back... Silly me.
If you’ve had access, my guess is a mistake was made and you were not supposed to be on the magic list.
Very cool.
Does this means that only the newest version of the Apple TV 4K can run this? I have a old 4K, what does it mean for it?
I do this ... but because I have no idea what that phrase was, not because I'm singing along.
How does the mic situation work? Will you need a USB mic? Can you use a BT headset?
I was skeptical but it works surprisingly well.
¹ http://preserve.mactech.com/articles/mactech/Vol.17/17.01/20... ² https://en.wikipedia.org/wiki/Carbon_(API)