This cool. But are these audio files transcribed, or just provided?
I downloaded a couple here:
http://www.repository.voxforge1.org/downloads/SpeechCorpus/T...
and it didn't seem to have a log of words labeled each by timestamp offset into the audio recording - which is the vital part for training a recognizer. Am I missing something?