Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
snakers41
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
snakers41
4y ago
One of Silero-VAD developers here. Check out our VAD here - https://github.com/snakers4/silero-vad
2.
▲
by
snakers41
4y ago
Neither of these
3.
▲
by
snakers41
4y ago
> Where the GPL was vague and exploited, the AGPL clarifies and closes the loophole It can be easily surpassed. Just create a simple wrapper and publish it. And voila, you are can use everything for free again) All of these FOSS licenses
4.
▲
by
snakers41
4y ago
It is PyTorch
5.
▲
by
snakers41
4y ago
Yeah, and instead of contacting the authors directly, you are making a public case about how unfair the NC license is.
6.
▲
by
snakers41
4y ago
> then is the intent of this project to only service corporations? One of the intents is NOT to service corporations for free or promote such services.
7.
▲
by
snakers41
4y ago
Oh, the same rhetoric used in depth about GNU AGPL licenses as well. And so nice to read the opinions of people explaining why a corporation X is not breaking your AGPL license and can use everything for free) The reality is much simpler -
8.
▲
by
snakers41
4y ago
It is possible, albeit with a significant simplification of the capabilities of the models (i.e. all of the SSML stuff will be left out). Also ONNX boasts native quantization that just works. But we are not currently actively working on thi
9.
▲
by
snakers41
4y ago
You can generate any sentence you want with provided colab examples.
10.
▲
by
snakers41
4y ago
See the links here https://www.reddit.com/r/MachineLearning/comments/v9rigf/p_s...
11.
▲
by
snakers41
4y ago
For STT models - yes. For TTS models - not yet.
12.
▲
by
snakers41
4y ago
What I hear is that the affluent "customer" does not respect licenses and you have no resources to enforce your license. Well ... maybe there is a correlation?
13.
▲
by
snakers41
4y ago
Many thanks for a detailed and thoughtful comment. > it is very fast and scales quite nicely on CPU with 4 threads (~ twice the speed), but not further (I tried it on a 64 cores box). Well, practically it does NOT scale even past 6 threa
14.
▲
by
snakers41
4y ago
HN removes the HTML anchors. The link (which I also copied below in the comments, b/c it was noticed after publishing) should be: - https://github.com/snakers4/silero-models#text-to-speech This link leads to the T
15.
▲
by
snakers41
4y ago
> Companies do steal software. They do not. The do not care about them as well. > Harvard Medical School is pirating software I wrote. > It's not worth a law suit against a $40B entity. > It doesn't matter what license
16.
▲
by
snakers41
4y ago
Our model can be simplified to remove all of the Python bits, and made to work with plain PyTorch jit-models or ONNX models (which both have a JAVA API), but we did not invest time in this yet. Typically, JAVA ~ commercial usage, and they a
17.
▲
by
snakers41
4y ago
We used to have AGPLv3 or similar, but we decided to abandon it for the reasons I explained in this (or above) thread.
18.
▲
by
snakers41
4y ago
> Do you mean you run your software on computers that can't run bash? It is explicitly stated, that PyTorch is the only real requirement. Bash is not required, i.e. models can be run on Windows or ARM with PyTorch. > Also, I'
19.
▲
by
snakers41
4y ago
These TTS models are not related to Kaldi, they are based off PyTorch and TorchScript. There can be made a simplified version, with ONNX models (or plain Torch jit) maybe and some outer logic, but we did not do it yet for lack of incentive.
20.
▲
by
snakers41
4y ago
This not only does not require a GPU, but also works on 1-4 CPU threads (!): - 8 kHz, 1 thread 15-25, 4 threads 30 - 60 - 24 kHz, 1 thread 10, 4 threads 15 - 20 - 48 kHz, 1 thread 5, 4 threads 10 the numbers are seconds of audio generated p
21.
▲
by
snakers41
4y ago
We already had our issues with local corporations neglecting the license (and being in general disrespectful towards the community), so we had to change it to CC BY-NC-SA to avoid this in future.
22.
▲
by
snakers41
4y ago
I am not sure, what can be more simple than 1 LOC invocation + minimal imports. It is true that the model is based on PyTorch + python, but the majority of complexity (like SSML parsing) is tucked inside of the model. Theoretically one can
23.
▲
by
snakers41
4y ago
Silero TTS works fast even on one CPU thread, this is the point
24.
▲
by
snakers41
4y ago
Also also, HN cuts the anchors in ULRs, so here is the full URL - https://github.com/snakers4/silero-models#text-to-speech
25.
▲
by
snakers41
4y ago
Also, a bit more detailed info here with voice samples: - https://www.reddit.com/r/MachineLearning/comments/v9rigf/p_s...
26.
▲
by
snakers41
5y ago
https://user-images.githubusercontent.com/36505480/145563071...
27.
▲
by
snakers41
5y ago
Stellar quality. Highly portable. No strings attached. Supports 8 kHz and 16 kHz. Models < one megabyte in size. Supports 30, 60 and 100 ms chunks. Trained on 100+ languages, generalizes well. One chunk ~ 1ms on a single thread. ONNX up
28.
▲
by
snakers41
5y ago
Stellar quality. Highly portable. No strings attached. Supports 8 kHz and 16 kHz. Models < one megabyte in size. Supports 30, 60 and 100 ms chunks. Trained on 100+ languages, generalizes well. One chunk ~ 1ms on a single thread. ONNX up
29.
▲
by
snakers41
5y ago
We have just updated our public voice detector, now it boasts features like: - Epic metrics (best in class as far as we know); - Much shorter and flexible chunk size (30, 60, 100 ms); - Radically faster inference, i.e. model forward pass t
30.
▲
by
snakers41
5y ago
Fast STT models rivaling Google - https://github.com/snakers4/silero-models#speech-to-text - https://github.com/snakers4/silero-models/wiki/Quality-Bench... - https://habr.com
More ›