Yes. In the video there is a short snippet where the audio is shown as a Fourier transformed image on the screen and a user is annotating the image of the sound using red boxes. This is a part of the process to train the ML model to recognize chainsaw sounds vs. other sounds.