3 ms·
> o1-preview is about on par with Anthropic's Claude 3.5 Sonnet in terms of accuracy but takes about 10X longer to achieve similar results to Sonnet. > o1's pe
by ckcheng 2y ago
> o1-preview is about on par with Anthropic's Claude 3.5 Sonnet in terms of accuracy but takes about 10X longer to achieve similar results to Sonnet.
> o1's performance increase did come with a time cost. It took 70 hours on the 400 public tasks compared to only 30 minutes for GPT-4o and Claude 3.5 Sonnet.
See details at https://arcprize.org/blog/openai-o1-results-arc-prize https://arcprize.org/blog/openai-o1-results-arc-prize
Also discussed here: https://news.ycombinator.com/item?id=41535694 https://news.ycombinator.com/item?id=41535694
- deleted 2y ago[deleted]