Gemini 2.5 Pro Capable of Winning Gold at IMO 2025
- ReLMLRM
Main:16 Pages
1 Figures
Bibliography:1 Pages
Appendix:51 Pages
Abstract
The International Mathematical Olympiad (IMO) poses uniquely challenging problems requiring deep insight, creativity, and formal reasoning. While Large Language Models (LLMs) perform well on mathematical benchmarks like AIME, they struggle with Olympiad-level tasks. We use Google's Gemini 2.5 Pro on the newly released IMO 2025 problems, avoiding data contamination. With pipeline design and prompt engineering, 5 (out of 6) problems are solved correctly (up to a caveat discussed below), highlighting the importance of finding the optimal way of using powerful models.
View on arXivComments on this paper
