Huh.. this is actually something I'm working on at the moment. The value is being able to asynchronously listen to audio of an article while keeping going with your work.
At the moment I'm relying on AWS primarily, it's got a couple good neural voices that I enjoy listening to, and then sync it up with S3 and a possible SNS (simple notification service). Glad to see someone else has seen a need for it, but I've also been thinking of how to do it agnostic of AWS.
It's possible at the moment for me to go into reader view, copy and paste the content into AWS polly interface, paste the bucket name, paste the SNS ARN, and then wait for it to finish, find it in the bucket and then open it.
I want that all in 3 steps, Start the Conversion, Find it easier in a better user interface, Hit Play.
And then from there, start implementing an SSML builder to modify the speed and prose of different paragraphs and stuff, but that's super far down the line.