4 ms·
OpenAI funded independent math benchmark before setting record with o3
- andrepd 2y ago> They also made a verbal agreement with OpenAI that prohibits the company from using the materials to train their models Hilarious.
- Frederation 2y agoCant trust anyone, ever.
- plsbenice34 2y agoI think it's more useful to assume everyone else is less blatantly deceptive than OpenAI, at least, so there's some hierarchy of trustworthiness to help lead you to the truth.
- aithrowawaycomm 2y agoElliot Glazer seems to have been caught in a contradiction: https://xcancel.com/ElliotGlazer/status/1880809468616950187 https://xcancel.com/ElliotGlazer/status/1880809468616950187 Here he says that Epoch is "developing" a private test set that OpenAI doesn't have access to, but elsewhere Epoch strongly implied that this already existed. This kind of makes me lean towards "Epoch AI lied" instead of "Epoch AI got played." (Even the coauthors weren't informed about the funding, so Epoch does not deserve a presumption of good faith.) I guess the real question: o3 was able to solve 25% of Frontier problems, so were these the problems whose solutions OpenAI had access to? If so, then that score is meaningless and dishonest.
- nioj 2y agoRelated https://news.ycombinator.com/item?id=42763231 https://news.ycombinator.com/item?id=42763231
- ChrisArchitect 2y ago[dupe] Discussion on source: https://news.ycombinator.com/item?id=42763231 https://news.ycombinator.com/item?id=42763231
- deleted 2y ago[deleted]