logoalt Hacker News

HenryNdubuakuyesterday at 9:40 PM1 replyview on HN

Users often stack a transcription model on top to get the voice prompt, then decode to actions. Think of Alexa and Siri.


Replies

dofmyesterday at 9:59 PM

Thank you.