4 ms·
Is each frame looked at separately? Given what is shown there seems to be no memory building context and pruning the options. Is that really hard to add?
by KayEss 10y ago
Is each frame looked at separately? Given what is shown there seems to be no memory building context and pruning the options. Is that really hard to add?
- omginternets 10y agoThere's something called "attentional neural networks" that attempt to do this. They tend to do very well in reading natural language, IIRC, but I've also seen them applied to video.