4 ms·
Why wouldn’t training small recurrent networks with GA scale? Genetics algorithms are embarrassingly parallel, even more so than training neural networks by gra
by joefourier 4y ago
Why wouldn’t training small recurrent networks with GA scale? Genetics algorithms are embarrassingly parallel, even more so than training neural networks by gradient descent. It’s not efficient sure but it’s trivially scalable.
- mark_l_watson 4y agoBecause training is not steepest descent.