5 ms·
The default configuration is cloud based. However you do get to choose where your voice data goes. They also have a method to keep the TTS local and I haven't c
by tekchip 7y ago
The default configuration is cloud based. However you do get to choose where your voice data goes. They also have a method to keep the TTS local and I haven't checked in on the project in a while but they were either working on or already had a local home server so you can make the whole system internal if you wish. I'd provide reference links but as noted site is down currently.
Edit: There are also several skills available that can point at full, local, downloads of wikipedia and the like. So if you prefer faster results and keeping some of your queries internal as a sort of hybrid thing that's an option as well.
- StavrosK 7y agoSnips were acquired? That's bad news, they were the only ones doing offline voice recognition...
- ocdtrekkie 7y agoThey were definitely not the only ones, just the ones with the most aggressively spammy marketing team. HN mods actually stepped in once or twice about it. Rhasspy is a really cool project that is compatible with a bunch of fully offline speech frameworks. PicoVoice is a super neat company too.
- StavrosK 7y agoOh, thanks for that, I've been looking for a good offline speech framework. Unfortunately, the hardest part for me is the microphone array, hopefully I can find something good there.
- ocdtrekkie 7y agoI haven't had time to play with Rhasspy, but the fact that you can plug and play components from different speech platforms is really promising, particularly when I want to interface it with my own assistant at some point. Wake word detection, speech to text, intent recognition, etc. is all split out and able to be plugged into separately.
- lifty 7y agoThanks for the tip! Both Rhasspy and PicoVoice look great. Anyone can recommend a microphone array that works with the raspberry pi?
- maddocche 7y agoLook for keystudio respeaker hat for raspberry. I got one for this exact purpose but still haven't got the chance to test it.
- lifty 7y agoVery cool. I see that respeaker has even a 6 mic array for the pi: https://respeaker.io/6_mic_array/ https://respeaker.io/6_mic_array/
- lukifer 7y agoSeconding both ReSpeaker and Rhasspy/voice2json, I've been having great luck with that combo. The docs and example code for ReSpeaker are great, very easy to work with.
- kelnos 7y agoI have the (older, discontinued) 6+1 USB mic array, but I'm in general wary of the ReSpeaker stuff. I have the original ReSpeaker Core v1 hardware as well, and it's incredibly unstable. It comes with a tiny amount of flash that you're supposed to augment with a SD card (which then gets overlayfs'd onto the root filesystem), but the hardware is so buggy that running things from the SD card causes random SIGSEGV and SIGILL all over the place (and yes, I've painstakingly verified that this is not a software issue). They essentially abandoned development of the v1 line (but appear to still be selling it!) for the v2 line, and I have a hard time trusting them. Admittedly, I haven't done that much with the mic array I have, mainly because I got burned out dumping tens of hours into trying to get ReSpeaker Core v1 working (which I ended up trashing; it was that bad). I'd like to use it with a Raspberry Pi; hopefully that works out.
- taf2 7y agoWhat about mozilla? https://github.com/mozilla/DeepSpeech https://github.com/mozilla/DeepSpeech
- bluGill 7y agoThat is one of the toolkits used by mycroft. It isn't everything needed for an assistant but if you want to make one it is probably the best starting point.
- lukifer 7y agoFor n=1, I got DeepSpeech 0.6 working on a Raspberry Pi, and the recognition accuracy was atrocious. Haven't tried the new 0.7 branch, though.
- taf2 7y agoKeep in mind deepspeech is an implement of speech recognition algorithm... the results will largely depend on model used which typically require lots of manual labor