4 ms·
Oh, yeah, I'm definitely aware there's still a quality gap when compared to proprietary online TTS options--but I'm specifically interested in FLOSS+offline for
by follower 5y ago
Oh, yeah, I'm definitely aware there's still a quality gap when compared to proprietary online TTS options--but I'm specifically interested in FLOSS+offline for my purposes.
And Larynx is game-changingly ahead of the other FLOSS+offline options.
(Probably the highest quality voice is the one that's used for the demo video narration--which when I first heard it I had to skip to the end of the video to confirm it wasn't a live human. :) )
My (mostly uninformed) impression is that there's room for training tweaking/improvement given how young the project is. And there's also multiple stages to the generation process so presumably there's opportunities at each stage.
- 2Gkashmiri 5y agoyeah, https://www.youtube.com/watch?v=hBmhDf8cl0k https://www.youtube.com/watch?v=hBmhDf8cl0k at the end says southern female english https://rhasspy.github.io/larynx/#en-us_southern_english_female-glow_tts https://rhasspy.github.io/larynx/#en-us_southern_english_fem... but the sample is NOT like the video, maybe the samples are old. the video is a great example