2 ms·
it requires hundreds of hours of speech data to make any difference, at least for a context-free asr. and gpu-powered high-end hardware, and several days for tr
by ftreml 7y ago
it requires hundreds of hours of speech data to make any difference, at least for a context-free asr. and gpu-powered high-end hardware, and several days for training.
training an asr model is totally different requirement than using it as a client