How Well Can You Hear Audio Quality? (2015)
npr.org
npr.org
320 kbps mp3 vs lossless 16bit WAV (cd quality) is extremely difficult to hear the delta without training, but if you know what the source instruments are supposed to sound like, such as hats and you can focus on their frequency band, you can hear compression artifacts and comb filtering.
This course really changed how I take in music and sounds in general.
I think of how mpeg video cannot accurately portray looking at a rippling lake.. the sinusoidal motion does not have much in common with blocky mpeg.
The overtone movements become disassociated and start splitting up into disconnected sine wave blocks. Transients also become softened.
With MP3s this is obvious at the high end, while lows/mids sound passable. The compression algorithm has more frequency bins at the high end, so the damage is more obvious there.
There's also a loss of low-level detail. Reverb tails lose definition, and the music loses front-to-back depth.
That being said, age, tinnitus and amazing progress within the audio compression sciences have made it more difficult for me to separate uncompressed WAV and 320 kbps MP3, though the 320 kbps files in the article had noticable artifacts in most cases.
FWIW, it had the complete opposite effect on me. I stopped caring about things like flac or other lossless formats for general music listening.
If the only way to maybe sometimes pick out a 320kbps mp3 from a flac one was to be on my best headphones, in my quietest room, focusing all my attention... the difference doesn't matter for practical listening. My ears may just be garbage, but it made me skeptical of people who make bold claims about what minutia they can discern -- although magical claims are more prevalent in the audiophile world ("these cables sounds great!") than the actual audio production world.
Along similar lines, I can't tell the difference between a $250 microphone and a crazy expensive mic in anything but ideal conditions -- and even then, I need to hear them side by side.
Buying a Ferrari won’t make you Michael Schumacher, but he could probably lap your car round a track faster than you ever thought possible.
This is not my idea btw, I read it somewhere, just forgot where I read it. Googling a bit around “musician audio quality ‘fill in the gaps’” didn’t ring any bells. I’ve been reading up on acoustics and audio engineering over the past year trying to get my personal music project together, so it may have just been from a book as well.
https://www.izotope.com/en/learn/psychoacoustics-how-percept...
As someone that has high-end gear, I can absolutely tell the difference in most cases between 320kbps MP3 and lossless when switching between the two versions of the same song. It's hard to put a finger on, but it just "feels" different to my ears/perception. However, after listening to exclusively one or the other for a few minutes or not A/B switching the same song, I have a much harder time detecting the difference.
Quality of the master and quality of the equipment matter much more, IMHO.
But: while it's still available, it's about ten years old, and I have no idea if it still runs on modern hardware.
Fidelity is indeed swell, but how important is it really? A lot of the music I like best, I first heard on a crappy recording or a shitty sound-system. There were A LOT of hit songs recorded before FM came along, and yet that 10k AM bandwidth (e.g. when the Beatles were big) didn't keep people from buying copies by the millions. Gasp! How can that be?
How much of the time I'm actually listening to music I like do I actually care about the fidelity? If band A's crappy cover is available in FLAC, will I choose it because band B's brilliant cover is only 128K mp3? NO.
This eternal argument has little to do with music. I still like 'Louie Louie' by the Kingsmen (a recording that cost $36), never mind how shitty the studio was. I'll still choose it over a FLAC from the studio master of "Never Gonna Give You Up". I can hear around the 'flawed' repro. Two kinds of 'purism' at work here. Pure genius beats pure fidelity.
I can read a novel or poem on a smartphone, in a paperback, on a fourth-generation photocopy, or in fine letterpress printing on 100% cotton paper. The reading experience might be a bit more comfortable in the last case, but the author’s words are the same in all.
I like the Kingsmen’s “Louie Louie,” too. I played it for my grandson yesterday.
Call me a snob, but usually I'd rather listen to a totally different Band C's song in at least 320k mp3 (preferably FLAC) than either Band A or B's in less quality.
But I enjoy music on a technical level as well. I find it very difficult to enjoy bad mastering or audio quality. There is enough well-mastered and recorded music out there nowadays that I don't feel I need to be subjected to badly-recorded music.
There are some recordings where I really wish they were done better, but very few that are really badly recorded that I can enjoy, no matter how good the music itself might be.
I'd love to see a similar experiment with newer codecs like Opus.
https://en.wikipedia.org/wiki/Tom%27s_Diner#The_%22Mother_of...
I chose the 320 kb/s for Katy Perry but was vacillating between that and the actual WAV. The one I completely screwed up was Jay-Z where I picked 128 kb/s. However, I knew I was lost with so much electronic glitching and distortion. While I heard differences, I had no idea what was a compression artifact versus a synthesis/mixing artifact. I also had no prior listening experience for either of these two, so no preconception for how it should sound.
For the other recordings, I think I could hear the masking of the psychoaccoustic models. Essentially, in the WAV I heard a "fuller" mix with overlapping sounds, while in the compressed versions I heard dropouts where a dominant sound was present in what felt like isolation. For Suzanne Vega this wasn't so easy, but I felt her voice and the reverb were more clear and natural in the WAV.
This was just with my old Sony MDR-NC22 noise canceling ear buds driven by my Thinkpad audio jack. My idea of a hi-fi environment is a Yamaha receiver and reasonable tower speakers like my Polks.
There are a number of audio artifacts in MP3s that are much more noticeable after training. However, many of them become much less discernable with variable compression rates and lower compression (>256kbps). It's possible to train on those artifacts, but I've always avoided it since I don't want the "projectionist effect" of making standard recordings less enjoyable.
This is not the case for many audiophiles. They want to fix all the bugs!
If it's something you love, it may make it your enjoyment a bit less.
For me, it's why I don't watch my wife pop zits in the bathroom. It's nice to not know about those flaws. :)
Netflix is AMAZINGLY low quality, even on the higher-quality tiers, and working with compressed video was one of the factors that ultimately led to me cancelling my subscription - I'd rather pay (or yarr) for an actual copy rather than something that's almost more compression-artifacts than actual video.
Once you've been shown the effects of overly strong video compression it is really hard not to spot them everywhere.
DVDs are horrible quality though. Compression everywhere and low resolution. The streaming services that I've watched are all better than DVD, at least 720p.
Blu-ray is obviously the better option, but they're more expansive and not always available, especially when it comes to stuff that isn't produced in the U.S.
[1]: https://imgur.com/a/XSCw824
…
[2]: https://en.wikipedia.org/wiki/Dynamic_range_compression
Red Hot Chilli Peppers' Californication album on the other hand is a prime example of terribly mixed music all across.
The cost, of course, is the the clip is horribly audible in the final waveform.
I listened to the samples multiple times for each test. As I went through, I kind of used a strategy to skip to and compare the busiest parts of the songs for "muddiness". Still, it was always very tough to tell between 320kbps and WAV. Most of them were decided on a subjective "it just sounds slightly clearer and fuller". I think I had the easiest time with the vocals only track; can attribute that to being a vocally-focused listener.
Side note on that last point: Only recently have I noticed how little I heard of drums in my music listening. Picked up edrums recently and realized how lost I was at drum beats "ideas". For melodies I feel like I have a large library of ideas to draw from. On the flipside, now I notice drums a lot more. Hopeful that others can relate to this too?
I did a similar experiment to myself years ago with a MiniDisc player and found that 128kbps was indeed the sweet spot for myself. Especially when space mattered.
While I am sure that more expensive equipment or years of dedicated study can help me discern the difference, I don't feel like my lack of refinement hampers my enjoyment of music in any way. If anything, I have no interest in training myself how to notice peas under mattresses.
Live and let live. I'd like less judgement for being like The Matrix character Cypher:
“I know this steak doesn't exist. I know that when I put it in my mouth, the Matrix is telling my brain that it is juicy and delicious. After nine years, you know what I realize? Ignorance is bliss.”
Edit: clarity of point.
I was also listening on my work setup, which isn't ideal: Bose QC35s plugged into an O2 amp and D50 DAC. I'd like to think I would've gotten more right with some flatter headphones.
The main things I noticed when comparing were:
* low end was smoother on the bass
* clarity of high end hits ("pss/tss" sounds)
* some vague sense of better imaging, like the sound stage was better
I didn't notice much difference in the midrange or vocals, but that might be because of the gigantic V-shape of my headphones.
And that is not even taking into account that MP3 (despite being really old and not the most efficient at this point) is good at what it does, especially at 320k.
A similar thing happens with MPEG2. At a certain point, all the bits thrown at it make it indistinguishable from source. Where newer codec shine is when you start saying "what if we do this at 1000kbps" or "600kbps".
If so Chapter 14 Fidelity and Distortion in Radiotron (RDH4) can be a good place to start:
https://ia801602.us.archive.org/18/items/bitsavers_rcaRadiot...
>The values of total harmonic distortion to provide objectionable distortion are 2 % with a frequency range of 15000 c/s and 10.8% with a frequency range of 3750 c/s for music, and 3% and 12.8% respectively for speech, with a pentode.
Even if almost nobody is using pentodes any more, you could do worse.
The whole thing is 1400 pages, you can see what engineers were up to with their slide rules back then.
Chapter 19 if you really want to know about decibels . . .
― Brian Eno, A Year With Swollen Appendices
http://web.archive.org/web/20090303092253/http://radar.oreil...
Personally I've had a similar experience: hearing a song somewhere played on low-quality speakers and liking it, then later finding the high-quality version and not liking it as much.
I remember I had one friend that when CDs came out didn't like them, but found that if he recorded them onto S-VHS tape sounded best of all, so he ended up using S-VHS like cassettes at home. In theory could have got a portable but that's taking it a step too far.
For every case, I picked the one I feel subjectively the best. Interestingly, I still prefer the 128k version from the "Tom’s Diner" example, even after knowing the "correct answer". This is amusing because according to the descriptions, this song was chosen as the "benchmark" for the algorithm, which could have resulted in some overfitting ;)
A man sits on a weight bench, wondering if the equipment is defective because he can't seem to bench press 800 lb...
It's a big step between esoterics and "please don't use the cheap shitty DAC that was integrated into your monitor to satisfy a feature checklist".
The amp part will matter a lot the more impedance the cans have, and many of the better headphones are high(er)-impedance. Might as well spring for a headphone amp + DAC combo just to see how it sounds. It's not that expensive, FiiO has some for under 80 bucks, for example.
That's why usually good audio tests use ABX where it's not about seeing if you can identify what is best, but rather if you can identify a difference at all. See for instance: http://abx.digitalfeed.net/
EDIT: I actually just tried the test and... ended up doing the same thing you did. On the Neil Young track I could hear a definite difference in the high pitched violin-like background, but I ended up selecting the 128kbps track as the "highest quality", maybe because the compression made it sound smoother.
https://gizmodo.com/is-there-anyone-stupid-enough-to-believe...
[Edit: Fixed Link]
The Apple USB-C to 3.5mm doesn't break the bank, works with Windows fine and the headphone nerds seem to rate it
If you can't hear the difference at a low price point, that's not a curse, it's a blessing.
Yeah the AMP and the DAC.
The DAC chip in a monitor will be the absolute cheapest one they can source, and is probably receiving a lot of noise from the monitor itself as well. The good news is that you don't need to break the bank to get something decent; a budget of ~$100 will get you something a million times better than a screen could ever deliver.
The $400 Element II [1] from JDS Labs (makers of the legendary Objective-2, a DAC+Amp that was designed to give top-end performance at bottom-end prices) is very literally as good as it gets without snakeoil entering the equation. THD when under a 150-ohm load is 0.0008%. That's very good. This is what I'll be upgrading to if my setup at work ever dies.
My work rig is an Aune T1 [2] which retails for $250 but Massdrop does them for $100-$150 sometimes. Yeah it's a "tube" but that doesn't matter as much as myth would have you; today's tubes are pretty transparent unless you deliberately get something that isn't. THD is good however nowhere near JDS Labs-level of perfection.
The Fiio Q2 mk2 [3] can be found for as little as $100 and it's a solid performer, though I haven't personally used it, Fiio's reputation is impeccable.
[1] https://jdslabs.com/product/element-ii/
[2] https://www.amazon.com/Aune-Second-Generation-Amplifier-Deco...
Indeed, the probability of choosing all the best clips by chance is 0.13%, and the probability of never choosing the worst (also by chance) is 8.8%. Choosing almost all the best clips, with exactly one mistake, would have the probability of 1.6%, and choosing exactly one 128k mp3 has a 26% chance.
Disclaimer: I have not actually performed the test.
[Edit: I'd also been drinking tonight which raises my blood pressure and I can hear slight ringing in my ears triggered by bloodflow or viscosity change. I can try the test again tomorrow.]
The best indicator for sound qualify is how long I enjoy actively listening to music. With a good source and sound production I can literally do nothing other than listen to the music for hours and not fidget with or read anything. Same with club systems. If it's not good I won't really have a good time for long. The Alpha Dynacord at Stereo in Montreal was amazing the time I heard it with 45 vinyls being played.
Used $50 Audio Technica ATH-M20X on $90 Behringer UMC202HD.
Did audio production in the past (amateur level).
Failed Coldplay, and Neil Young, the latter I argue has such poor source quality that it's hard to tell a difference.
Tip: listen to the "air" of you want to spot the difference, that's the first stuff the compressors throw away. "Air" is the ambient sound of the room, the tail of the reverb, the stuff you hear when you don't hear anything if that makes sense.
On the flip side of the coin, my wife regularly listens to music just out of her phone speaker and claims to not notice the difference.
Everyone enjoys music in a different way, and I'm glad there are plenty of options!
I listen to Tidal through a USB DAC connected to a preamp -> amp -> respectable speakers.
And I can definitely tell the difference. But no one can tell the difference on laptop speakers, the idea is nonsensical, or worse, spreading the idea there is no difference between uncompressed audio and highly compressed audio.
That is the point, isn't it? The absolute vast majority in of people listen to music on anything but gold-plated pre-warmed high-end systems. They listen to music on phones, on portable speakers, in headphones, while driving, cooking, taking care of kids, working... For any practical purpose there is no difference between uncompressed and compressed audio.
Generally I can usually (depending on the song) pick out 320kbps MP3 vs WAV on reasonable headphones (ATH-M50x), but I don't have a pair to hand to test these samples.
For the others I picked at random one of the two that sounded the most similar. In no instance could I tell 320Kbps and WAV apart. Could have been just luck that I didn't pick 128Kbps more than once.
Hardware is Bose QC II on low noise cancellation, connected by wire to a MacBook Pro and ears, one of which can hear crickets, the other not.
The inclusion of the Neil Young got a smirk from me because of his past antics with promoting his “Pono” high-res music service. (2012-2017)
The Jay-Z is there for similar reasons, explained in the opening: he was promoting his own high-res music service. (Still theoretically extant but he’s gone back to streaming on Spotify.)
And the Suzanne Vega track is there because it was one of the three or four songs that was used as a benchmark for the development of the MP3 format; if there is any song that should sound good as a shitty low-res mp3, it is Tom’s Diner.
But what about the Coldplay? And the Katy Perry? And the Mozart - and maybe this particular recording of Mozart? Are any of these entwined with the history of lossy audio compression, or attempts to sell lossless compression to a wider market than audiophiles?
How Well Can You Hear Audio Quality? - https://news.ycombinator.com/item?id=9654758 - June 2015 (71 comments)
How Well Can You Hear Audio Quality? - https://news.ycombinator.com/item?id=9743877 - June 2015 (1 comment)
How Well Can You Hear Audio Quality? - https://news.ycombinator.com/item?id=9688095 - June 2015 (2 comments)
And I've got HD598s and was listening through a UMC1820 interface.
That said, I'll take lossless FLAC files for a download any day, simply because it's future-proof in a way compressed audio isn't. I'm quite happy to listen at 320kbps on Spotify most of the time, or transcode to 96kbps Opus any FLAC I have when I load it on my phone.
In the end that compressed great song will still beat the uncompressed lame one, but how noticeable compression gets also depends on what you do with the material.
This test result is a tribute to sound engineers who make music good on every device.
(my audiphile dad did 'train' me... made me listen to the highend stuff he's been building/crafting since I was little. I owe him disliking 128kbps on almost every setup ]: I'm fine with 320kbps for most things though)
In particular, the coldplay song is one I've listened to alot, and I was able to instantly tell the difference in the instrumentation.
Listening on Apple Airpods Max, audio is getting piped through a virtual audio device via Loopback.
You can see the same thing with high-quality JPEGs vs PNGs for photos. There's a decent size reduction, but to see the difference, I have to open the files in Photoshop, zoom in, and toggle which one is visible. If I compute the difference in Photoshop, it's usually only a few levels different, and never more than ~8. That's really hard to see.
Sadly, they're not even using a modern codec like Vorbis or Opus.
That said, while I doubt I can tell the difference between FLAC and high-quality Vorbis, once I learned to listen for distortions on the high end, it's a little hard to unhear. Same as looking for blocking artifacts in JPEGs or video.
Music is not white noise. It does not maximize the channel capacity of the audio spectrum. As a rule of thumb, FLAC compresses most anything a human might call "music" in about a 2:1 file size ratio.
This might of course be the IKEA effect, but I'm quite happy nevertheless with my flac collection.
I missed Neil Young and, surprisingly, the classical music, although I suppose that classical music is used extensively when testing codecs.
That said, maybe the headphones compresses the wav as well...
The equipment in front of me in the control room alone is worth around 90000 (SSL board, RME conversion, Neumann kh410s for monitoring and serious acoustic treatment) — nothing super crazy when it comes to studio setups but certainly in the 'above average' area of things, even in this profession and certainly when it comes to anything normies are gonna experience. My most expensive headphones (I own around 10) are Sennheisers for ~1.5k, powered by a Phonitor, which cost around the same.
To toot my horn, I have been fairly successful at this. I work with chart pop/alt/indie musicians on a regular basis. I am actually in the rare position to have made a good living off of this.
While working on a Mix I am intensely focused on listening, to the point where I will not notice when being talked to. I am listening to and passing judgement on over 100 individual sources of audio running at the same time. I will hear if one of the 16 background vocals tracks is slightly out of tune. I will hear if there's an unpleasant resonance on a guitar somewhere for a split second. Absolutely nobody, including the artist, will ever be as acoustically intimate with the whole thing as I am during this time (this is probably true for any notable mix engineer).
Considering all this, the highly intensive and focused process, the considerable amount of money carefully spent (in contrast to audiophiles just throwing money at something because of what I would imagine are insecurities and boredom, I have to make a profit and thus spent my money carefully even) and my experience, I rank the difference between high quality cbr/vbr mp3 and lossless as so fucking negligible, that I will randomly chose either when making exports to compare my future working status to.
To elaborate on what this means: I will make an export, place it in the project, continue working on it, and then at a later point solo (= switch to) to those exports to check if I am going in the right direction, if I improved on the critical parts, if something got lost on the way.
To be absolutely clear on this: I would cost me an extra half second to make sure I always have a 24/96khz export of that mix as my reference instead of a maybe-i-dont-care mp3. I do not invest that half second. I don't know how to be any clearer than this.
Do not confuse this with me not being able to probably spot the difference. It's just that it is about a skill as spotting a difference between two apparently white walls. I am sure with enough effort and attention to the task anyone would be able to make their brain tune in on the the differences. But why the fuck would anyone do that? There is no upside. No, it will not make you a better music listener or the music more enjoyable. If anything you are doing more focusing on the wrong things.
Every now and then I stumble across this topic online and what used to be insecurity has turned into perplexion. I have no hope for or illusions about that the people who have trained their brains to search for mp3 artefacts (well done, I suppose) but maybe I can reassure some who seem to be on the fence on this topic: Do not dwell. This is a great way to waste time and money and get absolutely nothing of value in return. Aim for headphones you enjoy listening on. They are probably not that expensive. Make sure they have replacement parts available, specially for the ear pieces and cable.
Difference is too hard to notice in voice only audio example because it does not have that variation/detail in it anyway.
Only upon using a dedicated headphone amplifier with a moderately good pair of headphones should the answers become easy.
Seems to me like the experiment has value.
https://hometheaterreview.com/why-do-audiophiles-fear-abx-te...
Go to any audio store with speaker rows and listen to a few. They are hugely different, and none of them are cheap quality.
Which basically means that we are far from solving the accurate speaker problem.
The problem that we are far from solving is the problem of people buying inaccurate speakers.
128kbps on nice headphones will always sound better than WAV on garbage headphones.
Can we verify that the website choices are in fact legit? i.e. the 320k mp3 is in fact 320k and the uncompressed wav is in fact that. And also, that the browser isn't messing with the choices either (by e.g. caching the first file you play and reusing it for the other choices).
Also, I ordered an optical cable so I can use my amplifier's DAC rather than the onboard soundcard DAC.