2 ms·
You mean any repeatable benchmark will be saturated. The problem is that there is a huge perverse incentive. The intelligence is in the training layer not in t
by imtringued 1mo ago
You mean any repeatable benchmark will be saturated.
The problem is that there is a huge perverse incentive. The intelligence is in the training layer not in the model parameters, but the intelligence is really good at remembering things, so if you let it take the test, it can RL it.
- mikert89 1mo agothis is a short term problem, over 20 years benchmark gaming will be a blip