One Model Family, Two Gold-Level Results: Fine-Tuning Nemotron for IOI and IMO
NVIDIA fine‑tuned its Nemotron‑3 foundation model to achieve gold‑medal level performance in both the International Olympiad in Informatics and the International Mathematical Olympiad 2026, using a reusable recipe of domain data curation, supervised fine‑tuning, reinforcement learning and a generate‑evaluate‑refine inference loop. The results demonstrate that a single model family can be specialized to excel in both algorithmic coding contests and rigorous mathematical proof tasks, marking a significant step toward versatile AI specialists.