Research1 min read
Nemotron-IMO: Training for Olympiad Mathematics Problems
Researchers trained Nemotron 3 Ultra checkpoints using supervised fine-tuning and reinforcement learning to generate and verify proofs for olympiad mathematics problems. The resulting system achieved a gold medal score of 30/42 at the 2026 IMO, releasing the model, data, and benchmark for further research.
From arXiv cs.AI