NVIDIA Nemotron-3 models achieve gold-medal and top-human scores in IOI 2026
Researchers have developed a post-training pipeline for large language models that enables them to achieve gold-medal performance in competitive programming. By combining supervised fine-tuning, reinforcement learning, and a new feedback-driven strategy called GenCorrect, the system outperforms top human contestants on the IOI 2026 problem set.