Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
toebee
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
toebee
1y ago
Thank you for the kind words! We only support English at the moment.. Hope to add more languages in the future.
32.
▲
by
toebee
1y ago
Thanks you!! We personally used Quickpod and Runpod the most. But you can try it now on HF Spaces without spinning up GPUs yourself! https://huggingface.co/spaces/nari-labs/Dia-1.6B
33.
▲
by
toebee
1y ago
Thanks for the interest! We also enjoyed using E5-F2 :) You can try it now on HF Spaces: https://huggingface.co/spaces/nari-labs/Dia-1.6B
34.
▲
by
toebee
1y ago
Thank you so much for the kind words :) We only support English at the moment, hopefully can do more languages in the future. We are planning to release a technical report on some of the details, so stay tuned for that!
35.
▲
by
toebee
1y ago
We will try to make it work, but not sure if will be an easy task. For now, you can try with https://huggingface.co/spaces/nari-labs/Dia-1.6B
36.
▲
by
toebee
1y ago
Thank you! You can add audio prompts of calm voices to make them a bit smoother. https://huggingface.co/spaces/nari-labs/Dia-1.6B you can try it here!
37.
▲
by
toebee
1y ago
Thank you!! Works for English only unfortunately :((
38.
▲
by
toebee
1y ago
Thanks for the kind words! We're just following our interests and staying upwind.
39.
▲
by
toebee
1y ago
Thanks for the kind words! You can try it now on https://huggingface.co/spaces/nari-labs/Dia-1.6B Also, we'll try to update the Demo Page to something lighter when we have time. Thanks for the feedback :))
40.
▲
by
toebee
1y ago
Yes! But you would need to put together a LLM system that created scripts from the book content. There is an open source project called OpenNotebookLM ( https://github.com/gabrielchua/open-notebooklm ) that does somethin
41.
▲
by
toebee
1y ago
We're adding guides for Zero-shot voice cloning. You can try it using the second example on Gradio: https://huggingface.co/spaces/nari-labs/Dia-1.6B
42.
▲
by
toebee
1y ago
We just clarified in the README, sorry for the confusion ;( Note that the model was not fine-tuned on a specific voice. Hence, you will get different voices every time you run the model. You can keep speaker consistency by either adding an
43.
▲
by
toebee
1y ago
Only two voices at the moment... We will need to upgrade the dataset to make that happen, and are considering that as one of the next steps.
44.
▲
by
toebee
1y ago
https://huggingface.co/spaces/nari-labs/Dia-1.6B we fixed it!
45.
▲
by
toebee
1y ago
You can try it now on https://huggingface.co/spaces/nari-labs/Dia-1.6B !!!
46.
▲
by
toebee
1y ago
Thanks for the kind words :)))
47.
▲
by
toebee
1y ago
We'll try to give a high-level overview when we publish the technical report!
48.
▲
by
toebee
1y ago
Thank you for the kind words! We don't have plans for that yet, but you can always open an issue or RP on Github.
49.
▲
by
toebee
1y ago
We have a ZeroGPU Space provided by HuggingFace up and running! Test it now on https://huggingface.co/spaces/nari-labs/Dia-1.6B
50.
▲
by
toebee
1y ago
We're still experimenting, so do not have samples yet from the larger model. All we have is Dia-1.6B at the moment.
51.
▲
by
toebee
1y ago
Thank you for the contribution! We'll be merging PRs and cleaning code up very soon :)
52.
▲
by
toebee
1y ago
Sorry for the confusion. the license is plain Apache 2.0, and we changed the wording to "intended for research and educational use." The point was, users are free to use it for their use cases, just don't do shady stuff with
53.
▲
by
toebee
1y ago
We are in the progress of fixing it! Thanks for letting us know :)
54.
▲
by
toebee
1y ago
We use descript audio codec! I’m not sure if DAC works on iOS…
55.
▲
by
toebee
1y ago
Thank you for the kind words! Dia wasn’t fine tuned on certain speaker, so you will get random voices every time you run it, unless you add a prompt / fix the seed. The outputs are a bit unstable, might need to add cleaner training dat
56.
▲
by
toebee
1y ago
Thank you!! Indeed the script was inspired from a scene in the Office.
57.
▲
by
toebee
1y ago
Thank you for the kind words <3
58.
▲
by
toebee
1y ago
It is way past bedtime here, will be getting back to comments after a few hours of sleep! Thanks for all the kind words and feedback
59.
▲
by
toebee
1y ago
You can add an audio prompt and prepend text corresponding to it in the script. You can get a feel for it by trying the second example in the Gradio interface!
60.
▲
by
toebee
1y ago
Unfortunately yes at the moment
More ›