3 ms·
Anyone taking a single look at the ARC-AGI "challenges" can see things a 5 year old could reasonably solve. They're just jerking eachother off and sending each
by well_ackshually 22d ago
Anyone taking a single look at the ARC-AGI "challenges" can see things a 5 year old could reasonably solve.
They're just jerking eachother off and sending eachother the elevator back: "independent" ML engineer (worked at <large ML company> and currently runs <ML company looking to be bought out) writes a shitty benchmark (writes a single example and spams an LLM to make more variants) and releases it out as the BRAND NEW FRONTIER IN THINKING.
Every single benchmark has been catastrophically flawed and made by clowns.
- ewild 22d agoThe whole point is a 5 year old can solve it.
- vbarrielle 22d ago> Anyone taking a single look at the ARC-AGI "challenges" can see things a 5 year old could reasonably solve. Isn't that the goal of these challenges? Each release shows challenges that are very easy for humans, but are impossible for the models at the time of release (which demonstrates some missing generality). I think I've read the challenge authors say that, the day they cannot make a new challenge, then models are AGI.