5 ms·
Did anyone find a solution to have whisper differentiate between multiple speakers in a conversation and mark them in the written output?
by mariusio 4y ago
Did anyone find a solution to have whisper differentiate between multiple speakers in a conversation and mark them in the written output?
- freeqaz 4y agoYeah there are some models that I played with that can do this. They only work for 2 or 3 speakers currently though. They term for this is "diarization". https://huggingface.co/pyannote/speaker-diarization https://huggingface.co/pyannote/speaker-diarization
- swores 4y agoI wonder, do any conference call services (zoom, GMeet, etc) offer the ability to record each participant's audio stream separately in a way that would make it easy to transcribe them separately then combine?