3 ms·
Anyone using any reasonably good small speech to text os models?
by Johnny_Bonk 9mo ago
Anyone using any reasonably good small speech to text os models?
- garblegarble 9mo agoFor my inputs, whisper distil-large-v3.5 is the best. I tried Parakeet 0.6 v3 last night but it has higher error rates than I'd like (but it is fast...)
- Johnny_Bonk 9mo agoNice I'll try it, as of now for my personal stt workflow I use eleven labs api which is pretty generous but curious to play around with other options
- garblegarble 9mo agoI assume that will be better than whisper - I haven't benchmarked it against cloud models, the project I'm working on cannot send data out to cloud models
- BiraIgnacio 9mo agooh I've been looking into whisper and vosk in the last few days. I'll probably go with whisper (with whisper.cpp) but has anyone compared it to vosk models?
- woudsma 9mo agoI’m using whisper with superwhisper on my mac. I’ve assigned a key on my keyboard, when I press the key it starts listening and when I release it, the text gets copied to the current cursor location. It works pretty well.
- d4rkp4ttern 9mo agoParakeet V3 is near-instant transcription, and the slight accuracy drop relative to the slower/bigger Whisper models is immaterial when talking to AIs that can “read between the lines”.