2 ms·
In a traditional CNN you use the same convolution kernel across an entire image. Here you do the same but you convolve the features of a person (pixel) with the
by pfd1986 7y ago
In a traditional CNN you use the same convolution kernel across an entire image. Here you do the same but you convolve the features of a person (pixel) with the feature of its neighbors, but those change depending on the network. Drawbacks are: I) you add one new person to the network, the whole thing has to be retrained; ii) convolutions are more complex and expensive