ParentFull threadyorwba·Speech-to-text models predict the next token of text from the preceding tokens of text and the current tokens of speech.View on HN