Gemini 2.5 Pro Capable of Winning Gold at IMO 2025

21 July 2025

Yichen Huang

Lin F. Yang

ReLM

LRM

ArXiv (abs)PDF HTML Github (770★)

Main:16 Pages

1 Figures

Bibliography:1 Pages

Appendix:51 Pages

Abstract

The International Mathematical Olympiad (IMO) poses uniquely challenging problems requiring deep insight, creativity, and formal reasoning. While Large Language Models (LLMs) perform well on mathematical benchmarks like AIME, they struggle with Olympiad-level tasks. We use Google's Gemini 2.5 Pro on the newly released IMO 2025 problems, avoiding data contamination. With pipeline design and prompt engineering, 5 (out of 6) problems are solved correctly (up to a caveat discussed below), highlighting the importance of finding the optimal way of using powerful models.

View on arXiv

Comments on this paper