3 ms·
We have ImageNet experiments in Section 4 of the full paper: https://arxiv.org/abs/2010.08127 https://arxiv.org/abs/2010.08127
by preetum 6y ago
We have ImageNet experiments in Section 4 of the full paper: https://arxiv.org/abs/2010.08127 https://arxiv.org/abs/2010.08127
- choppaface 6y agoYou're not ablating anything there. What happens when Train Infinity (Train 150K) doubles in size (to Train Infinity_2 -> 300K)? What happens when you add an unseen class? These are real-world conditions that hamper existing theoretical estimation of the generalization gap-- the "ideal world" always gets larger. In Bengio's group paper (Predicting the Generalization Gap https://arxiv.org/pdf/1810.00113.pdf https://arxiv.org/pdf/1810.00113.pdf ) they actually do these sorts of ablations. Also, you use K (thousands) and $K$ (latex K) interchangeably; it's really hard to decipher is K is a variable or what you mean.