3 ms·
If you actually try current vision systems out on real raw video data as opposed to clean datasets of "good" photos pre-selected by humans, you'll see that they
by pakl 9y ago
If you actually try current vision systems out on real raw video data as opposed to clean datasets of "good" photos pre-selected by humans, you'll see that they are terribly far from human performance.
Same goes for translation systems.
Most current systems (by their very design) lack dynamical representation capability necessary for modeling interactions in/of the world. I hypothesize this is important for AI that actually gets what it's dealing with.