3 ms·
There's a lot of conjecture in this article. While I'm sure most voice assistant makers would love to be able to tag multiple voices simultaneously in realtime
by blixt 9y ago
There's a lot of conjecture in this article. While I'm sure most voice assistant makers would love to be able to tag multiple voices simultaneously in realtime and perhaps even in the background, it's simply not there in today's technology.†
The Alexa works in a crowded environment because it has directional microphones and facing the Alexa makes a big difference.
The Google Home assistant does differentiate requests based on voice pattern, but from my testing it actually only analyzes the word "Google". So person A can say "Okay", person B can say "Google", and person C can say "What's the weather?" and the Google Home will recognize person B (this might have changed since I did my testing, so don't quote me on that).
All this said, yeah of course any company can spy on you, and honestly Alexa is the least of your concerns. Laptops have 1-2 microphones built in. Phones have 3+ microphones. A lot of monitors and "smart" TVs come with microphones. Your intercom system can be wired into, and you can most likely hack into digitally based ones as well. Don't even get me started on IoT devices from non-IT companies.
A lot of malicious things can be done with technology today, but Alexa is most likely the safest device of the examples I gave. So yes – stay vigilant – but don't ignore the more obvious vulnerabilities in your home.
† There are some examples that prove we will be able to get there eventually (like Honda's ASIMO which has a demo of three people asking for different things at the same time), but nothing of the like has been seen in uncontrolled/noisy environments.
- nerpderp83 9y agoI have learned to never rely on something being, "technically not possible" that is only a side effect.
- blixt 9y agoMy point is not about whether or not it's possible because just about anything can be done eventually. It's simply a case of what is more likely: 1) Bleeding edge algorithms that can not only separate multiple speakers in parallel on a small device, but also transcode them and track their identities over periods of time and report back. 2) Alexa does exactly what it says on the bin because Amazon already extrapolated everything they need to know about you (including whether you have a teenage daughter who is pregnant[1]) from your last 3 text searches. Also see this relevant comic: https://xkcd.com/538/ https://xkcd.com/538/ Now I may sound defeatist but this is not my intention. Like I said you should stay vigilant, but this article is barking up the wrong tree and may in fact distract from the real dangers in mass surveillance and tracking. Those dangers are far more primitive, yet effective, than you might think. -- [1] https://www.forbes.com/sites/kashmirhill/2012/02/16/how-target-figured-out-a-teen-girl-was-pregnant-before-her-father-did/#7359c58a6668 https://www.forbes.com/sites/kashmirhill/2012/02/16/how-targ...
- nerpderp83 9y agoAwesome, thanks for the clarification. We tots agree. This article could be correct in 1 or 2 revs of all these devices. Which makes sense in our superscalar universe.