3 ms·
I think it would be good if they provided some samples on the readme. It would be good for example if their list of languages/accents could be sampled [1] [1]
by bArray 2y ago
I think it would be good if they provided some samples on the readme. It would be good for example if their list of languages/accents could be sampled [1]
[1] https://github.com/espeak-ng/espeak-ng/blob/master/docs/languages.md https://github.com/espeak-ng/espeak-ng/blob/master/docs/lang...
> eSpeak NG uses a "formant synthesis" method. This allows many languages to be provided in a small size. The speech is clear, and can be used at high speeds, but is not as natural or smooth as larger synthesizers which are based on human speech recordings. It also supports Klatt formant synthesis, and the ability to use MBROLA as backend speech synthesizer.
I've been using eSpeak for many years now. It's superb for resource constrained systems.
I always wondered whether it would be possible to have a semi-context aware, but not neural network, approach.
I quite like the sound of Mimic 3, but it seems to be mostly abandoned: https://github.com/MycroftAI/mimic3 https://github.com/MycroftAI/mimic3
- follower 2y agoFYI re: Mimic 3: the main developer Michael Hansen (a.k.a synesthesiam) (who also previously developed Larynx TTS) now develops Piper TTS (https://github.com/rhasspy/piper https://github.com/rhasspy/piper) which is essentially a "successor" to the earlier projects. IIUC ongoing development of Piper TTS is now financially supported by the recently announced Open Home Foundation (which is great news as IMO synesthesiam has almost single-handed revolutionized the quality level--in terms of naturalness/realism--of FLOSS TTS over the past few years and it would be a real loss if financial considerations stalled continued development): https://www.openhomefoundation.org/projects/ https://www.openhomefoundation.org/projects/ (Ok, on re-reading OHF is more generally funding development of Rhasspy of which Piper TTS is one component.)