3 ms·
Yes, this is essentially how AlphaGo and AlphaZero algorithms work to train superhuman Go/chess/shogi agents. It’s an elegant algorithm that is analogous to how
by willmarch 11d ago
Yes, this is essentially how AlphaGo and AlphaZero algorithms work to train superhuman Go/chess/shogi agents. It’s an elegant algorithm that is analogous to how humans learn games.
- zug_zug 10d agoWell except AlphaZero played 44 million chess games in that time (and actually played with a 44 core computer). So I'd like to point out that the human is still just a few orders of magnitude more efficient.
- willmarch 10d agoYes, we all know that biological systems are more efficient than machines through billions of years of evolution and natural selection but the overall process is largely the same (interacting with an environment, learning from results, improving underlying architecture, etc); efficiencies will come with more time and improvements.
- zug_zug 10d agoWell if we make AI that learns at the rates humans do, it'll fundamentally undermine and destroy the relevance of all existing AI. It sounds to me like you're saying "we basically are there it's just a matter of degree" and I'm saying "No it's orders of magnitude off and probably won't be using LLMs at all and maybe a very fundamentally different type of neural net technology that hasn't been invented yet."
- willmarch 8d agoI guess we'll find out soon enough...