Hacker News new | ask | show | jobs
by yorwba 10 days ago
Speech-to-text models predict the next token of text from the preceding tokens of text and the current tokens of speech.
1 comments

Thanks, I did some learning and it fell more into place.