8 ms·
I think when somebody trains code golfing LLMs with reinforcement learning they will inadvertently be smarter
by noopprod 4mo ago
I think when somebody trains code golfing LLMs with reinforcement learning they will inadvertently be smarter
8 ms·