3 ms·
The authors mention that before publications they tested these questions on Gemini and GPT, so they have been available to the two biggest players already; they
by fph 8mo ago
The authors mention that before publications they tested these questions on Gemini and GPT, so they have been available to the two biggest players already; they have a head start.
- data_maan 8mo agoLooks like very sloppy research.
- pickleRick243 8mo agoI don't think it's that serious...it's an interesting experiment that assumes people will take it in good faith. The idea is also of course to attach the transcript log and how you prompted the LLM so that anyone can attempt to reproduce if they wish.
- data_maan 8mo agoIf you want to do this rigorously, you should run it as a competition like the guys at the AI-MO Prize are doing on Kaggle. That way you get all the necessary data. I still think this is bro science.
- yorwba 8mo agoIf this were a competition, some people would try hard to win it. But the goal here is exploration, not exploitation. Once the answers are revealed, it's unlikely a winner will be identified, but a bunch of mathematicians who tried prompting AI with the questions might learn something from the exercise.
- data_maan 8mo agoBut everything has been explored in other datasets already. If only a bunch of mathematicians learn something, why are so many people talking about this, why is the NY Times posting about this? This is the attention economy at its worst.