3 ms·
Their validation set consisted of 210 positive images. The test set consisted of 20 positive images. These are very small evaluation sets for deep learning. My
by argonaut 7y ago
Their validation set consisted of 210 positive images. The test set consisted of 20 positive images.
These are very small evaluation sets for deep learning. My point is the work is promising but should be viewed with healthy skepticism (by default).
I would really not read anything in particular into "Broad and international participation... ...sample." That's just a claim in a paper, it's not "the truth".
- feral 7y ago> These are very small evaluation sets for deep learning. Evaluation is a statistics question, and it doesn't matter that the deep learning model used is high capacity and needs a lot of training data. There's nothing inherently wrong with validating a complex model on a small amount of data. The paper has a section 4.2 that gives a statistical analysis. Granted, it'd be nicer if they had enough data to show statistically significant differences.