3 ms·
One usecase that isn't mentioned here is call transcription. There are a lot of usecases where recording your own call and having some automated way to turn it
by oli5679 3y ago
One usecase that isn't mentioned here is call transcription. There are a lot of usecases where recording your own call and having some automated way to turn it into a transcript, and possibly generating some structured summary is a big time-saver.
OpenAi's Whisper is the most accurate transcription model that I know of. The weights are open-sources, so it can be self-hosted, or you use the API. The downside is that you have to roll your own diarization (seperating the text between speaker A/B). I used pynote audio.
Paid services like fireflies, and transcription tools built into your call software, are much easier to use, but lead to some transcription quality dropoff.