Comment by well_ackshually

2 hours ago

Anyone taking a single look at the ARC-AGI "challenges" can see things a 5 year old could reasonably solve.

They're just jerking eachother off and sending eachother the elevator back: "independent" ML engineer (worked at <large ML company> and currently runs <ML company looking to be bought out) writes a shitty benchmark (writes a single example and spams an LLM to make more variants) and releases it out as the BRAND NEW FRONTIER IN THINKING.

Every single benchmark has been catastrophically flawed and made by clowns.