Use this Whisper derivative repo instead - one hour of audio gets transcribed within a minute or less on most GPUs - https://github.com/Vaibhavs10/insanely-fast-whisper
It's a shame faster-whisper never landed batch mode, as I think that's preventing folks from trying ctranslate2 more easily.
Repos like https://github.com/SYSTRAN/faster-whisper makes immediate sense on why it's faster than the original implementation, and lots of others do so by lowering quantization precision etc (and worse results).
but this one, it's not very clear how. Especially considering it's even much faster.