4 ms·
I see DeepMind is still playing around with RL + search algorithms, except now it looks like they're using an LLM to generate state candidates. I don't really
by dinobones 2y ago
I see DeepMind is still playing around with RL + search algorithms, except now it looks like they're using an LLM to generate state candidates.
I don't really find that this impressive. With enough compute you could just do n-of-10,000 LLM generations to "brute force" a difficult problem and you'll get there eventually.
- richard___ 2y agoSigh. Just wrong