> Two speech-to-text models—outperforming Whisper
On what metric? Also Whisper is no longer state of the art in accuracy, how does it compare to the others in this benchmark?
On what metric? Also Whisper is no longer state of the art in accuracy, how does it compare to the others in this benchmark?
Curious if there's a benchmark you trust most?