3 ms·
Their best reported results are a hybrid effort though, here is one of the authors of the paper describing how they used programs generated by the LLM to extrac
by caddemon 3y ago
Their best reported results are a hybrid effort though, here is one of the authors of the paper describing how they used programs generated by the LLM to extract their own insights that then refined future iterations of their workflow: https://x.com/matejbalog/status/1735331210140819938?s=20 https://x.com/matejbalog/status/1735331210140819938?s=20
It can work by itself too but it is unclear at a glance how well since the main focus of the paper is the new mathematical benchmarks they achieved, i.e. their best results. Will have to read the paper more closely to say anything with high confidence, but based on their summary I'd guess the human in the loop part was pretty important here.