2 ms·
STT models just turn audio into text, though; they don't interpret what the text means or figure out how to translate it into other representations like functio
by ameliaquining 1mo ago
STT models just turn audio into text, though; they don't interpret what the text means or figure out how to translate it into other representations like function calls. You would use a general-purpose model for that.