The frontier live voice models (especially GPT Voice/Live or whatever they call it this week) do not have the typical AI-tells. So I wonder about a few things:
Why is this?
Why couldn’t whatever went into post-training these voice models be applied to the text-gen models to get rid of AI smells?
Could we set up a pipeline that leverages the voice models to clean up text that reeks of AI ?