An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.10712v1 Announce Type: new Abstract: We study how model post-training and test-time inference design affect natural-language proof generation for hard olympiad mathematics. Starting from Nemotron 3 Ultra, we train two specialist checkpoints using supervised fine-tuning and reinforcement…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-12 04:00 · arXiv cs.AI
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics