11 ms·
> Generating humanlike speech requires the model to act as though it thinks while speaking, while making use of filler words to make the speech sound extremely
by 22c 3y ago
> Generating humanlike speech requires the model to act as though it thinks while speaking, while making use of filler words to make the speech sound extremely realistic.
This is impressive from a "generating humanlike conversation" standpoint, but its application in terms of a support call seems rather boneheaded to me.
I accept filler words when speaking to a human because many humans need time to think. An AI generated response does not. The filler words are extra tokens that waste time and energy.
I'd feel somewhat miffed if I found out I was speaking to an AI that was intentionally wasting my time in an effort to sound more human. I only need a TTS voice to sound more human so that I can better understand what it's saying, not so that I feel convinced that I'm talking to a real human.