3 ms·
Hi, I used the [WhisperSpeech](https://github.com/collabora/WhisperSpeech https://github.com/collabora/WhisperSpeech) model for the TTS part after I did some se
by jpcl 3y ago
Hi, I used the [WhisperSpeech](https://github.com/collabora/WhisperSpeech https://github.com/collabora/WhisperSpeech) model for the TTS part after I did some serious torch.compile optimizations to bring the latency down. The Whisper speech recognition and the LLM were optimized through TensorRT-LLM by Marcus and Vineet.
It's not perfect but I am still extremely proud of how it came out. :)
- renus 3y agoWhisperFusion is fully open-source - https://github.com/collabora/WhisperFusion https://github.com/collabora/WhisperFusion
- stiffler01 3y agoTried this on 4090 and the responsiveness and real-time communication it offers are truly impressive. It has significantly improved my overall experience, especially in scenarios where minimal delay is crucial. Compared to WhisperFusion, Rabbit R1 feels like it's stuck in the past, they could maybe use the OpenSource WhisperFusion.