3 ms·
Yep, there are many reimplementations. Here is a reimplementation that swaps out a neural net with a GBDT to address compute costs: https://github.com/cgreer/a
by ta_tunestub 4y ago
Yep, there are many reimplementations. Here is a reimplementation that swaps out a neural net with a GBDT to address compute costs:
https://github.com/cgreer/alpha-zero-boosted https://github.com/cgreer/alpha-zero-boosted
- woah 4y agoHow does the performance of this version compare?
- ArtWomb 4y agoImagine its perfect for Computer Backgammon, but overfits higher dimensional spaces ;)
- cgreerrun 4y agoDepends on game/environment and—since it's using a GBDT and not a NN—how good you are at feature extraction/selection for your problem. High level, I'd say it's a good way to test a new environment w/out spending time/effort on GPUs until you understand the problem well, and then you can switch to the time/money costly GPU world.