Floppycasts – 1.44MB Podcasts
ajroach42.com
ajroach42.com
https://auphonic.com/blog/2018/06/01/codec2-podcast-on-flopp...
At the top end of that, I think that's not the case. Certainly starting at 3200bps, Opus just becomes a non-option (and codec2 becomes an option).
I have a trove of librivox audiobooks in codec2 from some tests I did, and I've given some of them extensive listens to get a feel for it. Without an improved decoder that isn't a WaveNet, it's still not always fun trying to understand what is said.
LPCnet gives similar quality with much less compuation needed:
http://www.rowetel.com/wordpress/?p=6482
Can you elaborate or put some concrete numbers on this? :)
<bgmusic character="spooky" />
<archiveaudio year="1972" type="cop" tone="evasive" />
<foley type="tirescreech" />
<foley type="river" weight="300" />
Golly - you could make any story sound like it was a Radio 4 play
So I'd use my archiveaudio tag from the example and the algo would output the band-limited frequencies of what sounds vaguely like a cop with speech patterns signaling defensiveness and irritation.
It would all sound perfectly like a podcast if you were listening to it in a sleep state. Then if you concentrated on it you'd realize there's no coherent plot or even any actual natural language being spoken in the entire recording.
My Sansa MP3 player, Acer laptop, and gateway PC all had different ways of 'rendering' the Orchestra hit instrument in particular, which made for a rousing Termina Field listening experience on the bus or at Grandma's compared to the family computer.
"It is expected that an implementation will take some care in achieving this ideal, but it is reasonable to consider lower-quality, less-costly approaches on lower-end hardware."
Luckily nearly no one seems to be using this API so I doubt there really are anywhere near as many "approaches" as you can hear from MIDI file players.
[1] It is expected that an implementation will take some care in achieving this ideal, but it is reasonable to consider lower-quality, less-costly approaches on lower-end hardware.
Of course I'm only spitballing here, this would require some pretty advanced research to be made.
WebVTT, reproduced at the listener side by TTS? At some point, there's a quality chasm that can only be leaped by using a format other than sampled audio.
My question for you is: Why do we need that?
Lots of advanced research has been done in this area. These are the results as of now.
Things will hopefully continue to improve.
Do we though? In the current state of technology podcasts can be quite long and still consume less bandwidth than a HD movie trailer.
Something like Vocaloid parameters would be interesting, but the synthesis engine is probably pretty intensive.
(To be clear, I think this is an idea that's fun to think about as a novelty but probably a dead end otherwise.)
<ad type="wysiwyg-website-builder" />
<ad type="overpriced-mattress" />
But opus still sounds pretty good until you get down to around 10kbps (this, iirc, is 6 or 7.)
But then it means you can make your podcaster say anything you want.
Careful with your units there.
Fun fact, the spring rate for the slidy thing (technical term) is measured in lb m/in cm!
No satisfying chunk sound though :-P.
I love to imagine using my 4G connection to download like 60 hours of podcasts in the few minutes before leaving on vacation...
It all fascinates me. It is amazing to see us push the boundaries of communication under such restrictions
When I actually release, I'll do it at a bitrate that's more listenable, but still optimized for space.
I imagine it would be less of a size difference and more noticable at higher bitrates.
As for resampling, it depends on the specifics of the resampler used by a specific codec implementation. If a resampler is terrible, it could add artifacts like ringing and aliasing that would degrade the encoding quality.
A great souce for resampler comparisons: http://src.infinitewave.ca/
I prefer the sound of the MP3 in the examples, although both are really dull. I think using some modern codec like WaveNet or others, whilst sounding better, kind of defeats the point. I think MP3 is really as new a codec as I would want to go, especially as pretty much any device that can play audio that is actively used today can read an MP3 file. (I'm somebody who still uses an MP3 player.)
That said, I think these floppycasts should simply be shorter, the easier they are to create, the more likely the idea won't die in the crib. I think 15 or 20 minutes is really not too bad for a start (especially a solo act), specials could come in multiple parts and still be in keeping of the theme.
Opus file is tuned for speech, mp3 isn't.
The opus file sounds unbearable when the music is playing, but is clearer otherwise.
Double the bitrate and opus is way ahead.
I will say the music part was interesting. I could mostly make out the song on the MP3. On Opus, I could barely tell there was music there. A lot of it was just missing.
A big flaw in the OP's testing is that he converted the audio to 8-bit PCM before feeding it to the encoders. That added 48dB of high-frequency noise and distortion for absolutely no reason. He might have found his sibilants ("s" sounds) sounding better had he not done that.
I havent decided whether this is a serious proposition yet....
As long as the podcast doesn't approach MS Office 97 territory (50+ disks).
Meanwhile phones are always getting faster and faster and always becoming obsolete every year. Throwing software away means we are forced to throw hardware away.
I wish that someday we could just decide to stick to a single system and not change things, so that system can last 10 years and work on the same durable phone. It would require making hard choices in OS and software design, but it would be really worth it.
I hope a day will come where computers or smartphones will be able to last 10 or maybe 20 years and still be affordable. To be honest I don't think I want to be involved in learning how to develop apps on any phone for those reasons. At least microsoft and linux are able to maintain a minimal amount of backward compatibility. For developers that's really important. You don't have a good ecosystem or a good phone if you cannot attract developers to build apps.
Would you mind linking to the uncompressed source as well?
And also this all reminds of old demoscene/music tracker days when groups would release "music disks"