3 ms·
It supports XTTSv2 which is currently the open-weight state of the art. So, pretty damn good (https://huggingface.co/coqui/XTTS-v2/blob/main/samples/en_sample.w
by lelag 2y ago
It supports XTTSv2 which is currently the open-weight state of the art. So, pretty damn good (https://huggingface.co/coqui/XTTS-v2/blob/main/samples/en_sample.wav https://huggingface.co/coqui/XTTS-v2/blob/main/samples/en_sa...).
Too bad that the project is in limbo after Coqui (the company) folded. The license limits the use of the weights to non-commercial usage unless you buy a commercial license, and there's nobody left to sell you one now.
- qup 2y agois there anyone to sue you? (how does that work?)
- deleted 2y ago[deleted]
- NikkiA 2y agoHonestly, I don't think that sounds as human as piper does, but that's probably a function of the voice model files more than anything, fe, en_US 'amy' sounds artificial, but hfc_female sounds more realistic on the Piper samples. https://rhasspy.github.io/piper-samples/ https://rhasspy.github.io/piper-samples/