3 ms·
I don't get why this matters at all? I looked at the o1-preview paper, Putnam is not mentioned. Meaning that: a) OpenAI never claimed this model achieves X% on
by bary12 2y ago
I don't get why this matters at all? I looked at the o1-preview paper, Putnam is not mentioned. Meaning that:
a) OpenAI never claimed this model achieves X% on this dataset.
b) likely, OpenAI did not take measures to exclude this dataset from training.
Meaning the only conclusion we can draw from this result is: when prompted with questions that were verbarim in the dataset, performance increases dramatically. We already know this, and it doesn't say anything about the performance of the model on unseen problems.
- whimsicalism 2y agoyep, welcome to hn