3 ms·
This is not good reasoning. Humans need at least dozens if not hundreds of reinforcement sessions to only make legal moves, and still occasionally fail (conside
by hackinthebochs 19d ago
This is not good reasoning. Humans need at least dozens if not hundreds of reinforcement sessions to only make legal moves, and still occasionally fail (consider pins, discovered check, failing to respond to check). LLMs must one-shot a competent game after imbibing a mass of disconnected units of information about chess. Nothing about the two are similar.
See my comment here for more: https://news.ycombinator.com/item?id=49725306 https://news.ycombinator.com/item?id=49725306