3 ms·
> Two speech-to-text models—outperforming Whisper On what metric? Also Whisper is no longer state of the art in accuracy, how does it compare to the others in
by progbits 2y ago
> Two speech-to-text models—outperforming Whisper
On what metric? Also Whisper is no longer state of the art in accuracy, how does it compare to the others in this benchmark?
https://artificialanalysis.ai/speech-to-text https://artificialanalysis.ai/speech-to-text
- jeffharris 2y agoWe've been using the FLUERS eval and you can see comparisons to other models on the market in the post https://openai.com/index/introducing-our-next-generation-audio-models/ https://openai.com/index/introducing-our-next-generation-aud... Curious if there's a benchmark you trust most?
- lern_too_spel 2y agoFLUERS and GP's Common Voice dataset focus on read speech. I've observed models that perform well on these datasets be completely useless on other distributions, like whispered speech or shouted speech or conversational speech between humans who aren't talking to a computer.