Remix.run Logo
yorwba a day ago

Speech-to-text models predict the next token of text from the preceding tokens of text and the current tokens of speech.

hackrmn 14 hours ago | parent [-]

Thanks, I did some learning and it fell more into place.