4 ms·
there is no training code and devs don't plan on ever releasing it
by MitPitt 3y ago
there is no training code and devs don't plan on ever releasing it
- huggingmouth 3y agoIt's only a matter of time.
- woodson 3y agoIt’s mostly there in https://github.com/lucidrains/audiolm-pytorch#hierarchical-transformers https://github.com/lucidrains/audiolm-pytorch#hierarchical-t.... They just used FAIRs EnCodec (https://github.com/facebookresearch/encodec https://github.com/facebookresearch/encodec) instead of soundstream.
- dragonwriter 3y agoThe voices aren’t the model; while the model takes cobventional training for which code is not provided, voices are, or at least can be, built by what could be described as “accumulated in-context learning”. Every time you run text with a voice (which can be null) through the inference process, the result is an audio waveform and an updated history prompt.