news.volyx.in

OpenAI claims gold-medal performance at IMO 2025 (twitter.com)

479 points by Davidzheng · 375 days ago · 698 comments on HN

Article summary

OpenAI claims its experimental LLM has achieved gold-medal performance on the International Math Olympiad (IMO) 2025, solving 5 out of 6 problems under the same rules as human contestants. The model was evaluated on its ability to craft intricate, watertight arguments at the level of human mathematicians. The achievement is seen as a significant milestone in AI progress, demonstrating the model's capability in sustained creative thinking and general-purpose reinforcement learning. The model's solutions to the 2025 IMO problems are available on GitHub.

Main themes

  • AI progress
  • Mathematical reasoning
  • Information retrieval
  • Goalpost movement
  • Real-world applications
  • Transparency in AI research

What commenters say

  • Some commenters argue that the IMO problems are not particularly difficult and that the model's achievement is overhyped.
  • Others disagree, stating that the problems are challenging and require a deep understanding of mathematical concepts.
  • There is a debate about whether the model's performance constitutes true reasoning or just smart information retrieval.
  • A few commenters note that the model's ability to perform well on the IMO does not necessarily translate to real-world applications or more complex mathematical problems.
  • Some express concern that the goalposts for what is considered impressive in AI are constantly being moved.
  • Others see the achievement as a significant milestone in AI progress, demonstrating the potential for AI to make rapid advancements in various fields.
  • The discussion also touches on the importance of transparency and methodology in reporting AI competition performance results.
  • A few commenters mention that the velocity of AI progress is exceeding the velocity of goalposts, making it challenging to evaluate the significance of achievements like this one.