4 ms·
The short answer is that everything is streaming — as tokens come back from ChatGPT we send them as soon as possible to the synthesizer. The long answer is foun
by ajaynraj 4y ago
The short answer is that everything is streaming — as tokens come back from ChatGPT we send them as soon as possible to the synthesizer. The long answer is found in our code[0] :).
[0] https://github.com/vocodedev/vocode-python/blob/main/vocode/streaming/streaming_conversation.py https://github.com/vocodedev/vocode-python/blob/main/vocode/...
- famouswaffles 4y agohow is it sounding good though. usually text to speech models need the full context to sound reasonable.
- KianHooshmand 4y agoWe chunk it up per sentence so it has some context!