6 ms·
Decades ago, I was treated to an ABX test in my brother's recording studio. I easily recognized and preferred a 24/192 master he played versus the 16/44.1 down-
by saltcured 3mo ago
Decades ago, I was treated to an ABX test in my brother's recording studio. I easily recognized and preferred a 24/192 master he played versus the 16/44.1 down-mix. I honestly don't know whether there was something wrong with the down-mix, but qualitatively it did feel like it was "muffled" and coming from speakers, while the master really felt like live performance. He was surprised that I could tell them apart.
I also spent a lot of time ripping my old CDs to FLAC and trying different MP3 and AAC encoder settings to get playback that felt transparent enough to me. I could never tolerate Sirius/XM radio streaming due to the horrid compression I heard with every futile attempt. I still seem to have more sensitive hearing than most people around me, but in my 50s I know it isn't what it once was.
I never had huge budgets, but did strive for hi-fi in my limited ways. I used things like toslink and HDMI to send raw PCM data from Linux to my Yamaha A/V receiver's DACs + amplifier to drive somewhat nice Polk tower speakers. But then COVID-19 happened, and this stuff was packed up to move house.
Nowadays, music playback is streaming with mundane "subwoofer + satellite" PC speakers or MP3 playback with a mini-SD card permanently parked in my car's infotainment system.
- empiricus 3mo agoEven for PC, I recommend some cheap studio monitors.
- bob1029 3mo agohttps://www.bhphotovideo.com/c/product/964752-REG/yamaha_hs8_8_2_way_bi_amplified.html https://www.bhphotovideo.com/c/product/964752-REG/yamaha_hs8...
- bzzzt 3mo agoThose things have huge drivers and are probably too big for a lot of rooms. Unless you (and your neighbors) absolutely want to have that thumping sound and you go out of your way to kill unwanted bass reflection you're probably better off with the HS5 or something similar.
- saltcured 3mo agoYeah I'm just lazy about dealing with the room. If I find the motivation, I'll pull the original equipment back out of storage.
- vor_ 3mo ago> Decades ago, I was treated to an ABX test in my brother's recording studio. I easily recognized and preferred a 24/192 master he played versus the 16/44.1 down-mix. I honestly don't know whether there was something wrong with the down-mix, but qualitatively it did feel like it was "muffled" and coming from speakers, while the master really felt like live performance. He was surprised that I could tell them apart. As referenced in the article, a common explanation for those audible differences is that the high-resolution version of the album is sourced from a different master.
- TheOtherHobbes 3mo agoThis is easy to disprove by downsampling from a 24/192 source to 16/44.1 Even if the downsampling is (close to) ideal there are obvious differences. In fact if you can't hear the difference between 24/192 and 16/44.1 you shouldn't be working in audio. (Doesn't apply to consumers. Does apply to musicians and engineers.) It's like being colour blind. And if you don't understand the math behind quantisation, you shouldn't be posting pseudo-scientific videos where you use an oscilloscope and a cheap spectrum analyser - both tools with very limited resolution - to "prove" your point. 16 bit isn't enough for hard, objective reasons. One is that the noise spectrum of quantisation is not simple. Most people assume it's something close to plain white noise, but it really isn't. It's actually a very complex spectrum with some prominent peaks at specific subdivisions of the sample rate. Those frequency peaks are significantly above audibility. 24-bit quantisation shrinks them below audibility. The other is that most people can hear dither/noise-shaping at 16-bits. That adds a single bit of noise which should - if you're being very literal - be far below the threshold of audibility. But it clearly isn't. These two facts are related. The more complex reason is that listening is an active perceptual process. The brain does a huge amount of processing to separate sources and place them in a perceptual field which includes information about perceived object type, distance, and ambience cues. Some of those cues are very quiet, and we don't hear them linearly. So using sine waves as some kind of perceptual reference for audibility is nonsensical. We hear much more complex signals in an active way, and if there's information missing in the quiet parts - which there is with limited quantisation - then the signal simply isn't accurate.
- 3mo ago
- nullc 3mo agoThis is an extremely hard comparison to do well. I'll give a few examples as to why: Small differences in gain are ABX able much more readily than differences in noise at the 16 vs 24 bit level. So if the signal chain gives even a small difference in gain between the samples that's what you'll track. A reasonable conversion path to 16 bits for mastering will also apply dithering and some kind of brickwall limiting (you have to limit after the dither or as part of the dither as dither can change levels!), and this can result in gain changes. The DAC may behave differently or have outright bugs for some configurations too. This is particularly true wrt reconstruction filters for sample rate differences. And if you were comparing 44.1k and 192k then the physical DAC itself was likely running at a different rate and its _analog_ filters are probably better optimized for one vs the other (this is less true for 48k vs 192k, as the hardware likely runs at the same rate for both). So one answer to this comparison can be "on this particular hardware this rate is better than that rate"-- but that's a implementation property not a property of format choice. You might think, "okay I'll use a mathematically perfect down and up conversion process and run the DAC in the exact same configuration for all cases". But even then you run into issues like after reconstruction the _inter sample_ peak levels will be higher than the levels of the samples, so you have to handle that and in a way that doesn't produce a gain difference between the two configurations. (probably by running your perfect process and finding the gain level that results in no limiting, then making the gain of the original match). And then for the high rate vs non-high rate you have to deal with the fact that most amplifiers are not particularly linear (compared to well constructed software at least!) and that any real speaker is very far from linear. This means that the presence or absence of ultrasonics will change the audio in the 0-20khz band.. Before you think "well that could be a reason that high rate is better" observe that if there was some consistently good effect from the ultrasonics you could just bake it into the low rate sample. > but in my 50s I know Yeah if you're in your 50's you're absolutely not hearing differences way up above 20khz (especially if you're male), I bet you can't even hear CRT flybacks from 100 yards anymore. :P Most people have no idea how much their high frequency hearing degrades as they age because it plays approximately no role in your life, but it's real, dramatic, and as far as I know happens to everyone. I don't mean to discount your experience: I don't really doubt that it was real. But answering the general question of the necessity of low vs high rate probably takes a team of experts, armed with test gear and the designs of the HW/SW in question, to vet the test configuration. Testing a _particular_ configuration without the ability to distinguish its implementation quirks from format-fundamentals is much easier and that's what most attempts to test this question are actually testing. By testing in a recording studio you were doing far better than most such comparisons. Usually people try comparing different files and they're comparing entirely different mastering processes. Files made for the "high res" market will often have much less compression and limiting then files made for commercial radio play / casual listening... and truly do sound obviously much better. Some of my favorite recordings are rips from vinyl. Vinyl is an awful format from the perspective of audio fidelity, but it's also pretty intolerant of excessive compression and limiting because the record will skip if the needle is bouncing off the rails. And more recently I suppose they also avoid over compression there because of the difference in target listener/environment.
- Applejinx 3mo agoThat would be how you'd go about telling, sure enough. You can't go by 'frequencies' or distortions or anything like that, these analog departures from convincing reality aren't how digital failings manifest. You try to hear the brickwall by the muffled, enclosed quality and possibly by the weird pre-ring blurriness of the filter making things sound more vague than they have to be, and you hear the truncation not because it is audible 'distortion' as we know it, but because depth collapses and it sounds like it's coming from the speakers and not being a separate space behind/around the speakers. At no point will it be the most glaringly obvious thing but it'll never be 'distortions' as we imagine them, it's more a 'pod people' lack of personality thing. Like a much subtler version of listening to AI music :) I'm quite happy with 24/96 as suitable overkill for anything I might want to hear or do. Neil Young went hard on the proposition that 192 was necessary. Sold the Ponoplayer, I had one but it died on me, battery failed eventually. It really did sound awesome beyond just about any other listening device I've ever heard…
- TheOtherHobbes 3mo ago24 > 16 is not debatable. Sample rates are more complex because then higher the clock rate the more you get distortions from jitter and the design of the DAC/ADC. Most converters introduce different artefacts at different sample rates, especially at the prosumer end, so you're not comparing like for like. The last couple of generations of converters have gotten a lot better, so 192kHz today is likely to sound cleaner and smoother than it did ten years ago, where there was a good chance the clock was quite jittery. Personally I don't think it's worth the extra bandwidth for playback, but I can understand why some people might want it. Generally all of these "debates" come down to people who think math > circuitry. All real designs are imperfect trade-offs. They all have issues, and arguing as if converters are perfect when they never are, and the imperfections can be benched objectively, is... not very scientific.
- bzzzt 3mo ago>Generally all of these "debates" come down to people who think math > circuitry. All real designs are imperfect trade-offs. They all have issues, and arguing as if converters are perfect when they never are, and the imperfections can be benched objectively, is... not very scientific. There is one purely objective benchmark: a true blind test. You can believe if something is different or not, but if nobody's capably of hearing the difference, does it matter?
- amluto 3mo agoOne possibility (pure speculation) is a bad antialiasing filter. The Nyquist frequency at 44.1ksps is 22.05kHz, which is only ~10% above the audible band. This means that you need a rather sharp filter both when downmixing and when playing to avoid potentially audible aliasing into the audible band or attenuation within the audible band. If you look at a site like audiosciencereview.com and pull up measurements of a DAC or ADC, you can find graphs of the antialiasing filter response. Some are great and some are not. One could think of 16/44.1 PCM as being a codec that is potentially perfect but requiring some degree of care to encode and decode correctly.