3 ms·
I think o1 and o3 are OpenAI stretching the usefulness of 4o. Anecdotally the latest Claude model is naturally better than 4o, and having it use inference token
by ynniv 2y ago
I think o1 and o3 are OpenAI stretching the usefulness of 4o. Anecdotally the latest Claude model is naturally better than 4o, and having it use inference tokens to think through or verify work results in similar o1/o3 gains. Are we sure we can't do even better than Claude 3.5 Sonnet/new without inference tricks? We're only a few generations into LLMs... I don't see why not.