Transcription is arguably one of the must useful enterprise AI tools avaliable. But i sure as heck wouldn't trust the cloud with it.
Is it because of a globally trained model (as opposed to trained[tweaked on] on context specific data) or because of using different classes of models.
It could be they simply use a mediocre transcription model. Wispr is amazing but would hurt their pride to use a competitor.
But i feel it's more likley the experience is; GPT didn't actually improve on the raw transcription, just made it worse. Especially as any miss-transcipted words may trip it up and make it misunderstand while making the summary.
if i can choose between a potentially confused and misunderstood summary, and a badly spellchecked (flipped words) raw transcription, i would trust the latter.