It's not just voice, but conversational interfaces in general that have this discoverability problem. That means chatGPT. I'm glad people are discovering it, if you excuse the pun. CUIs are not a panacea, not a replacement for conventional UIs.
So the difference wrt discovery is that you only have to gesture at what you wanna do and, if a matching action exists, there is a chance it will be understood.
I'd wager we'll see a renaissance of voice assistants with LLMs, especially once the good-enough ones can run on device.
Not really conversations though, more like accessible commands. I don’t think we get conversations until the tech improves and the latency goes way down, meaning on device processing of speech at the very least.
Ow, right in the AARP flyer.