Your software is such a great idea. The UI could use some work to be more attractive, but the core functionality is top notch.
Here's a suggestion- using Speech.framework[0] you could probably quite easily transcribe the audio and identify filler words ("umm", "hmm", etc) and add an option to automatically exclude those as well.