Hacker Newsnew | past | comments | ask | show | jobs | submitlogin
OpenAI O1 Results on ARC Prize (twitter.com/arcprize)
9 points by rzk on Sept 13, 2024 | hide | past | favorite | 2 comments


> Results: both o1 models beat GPT-4o. And o1-preview is on par with Claude 3.5 Sonnet.


> o1-preview is about on par with Anthropic's Claude 3.5 Sonnet in terms of accuracy but takes about 10X longer to achieve similar results to Sonnet.

> o1's performance increase did come with a time cost. It took 70 hours on the 400 public tasks compared to only 30 minutes for GPT-4o and Claude 3.5 Sonnet.

See details at https://arcprize.org/blog/openai-o1-results-arc-prize

Also discussed here: https://news.ycombinator.com/item?id=41535694




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: