NVIDIA post-training pipeline hits gold-medal IOI performance, topping top humans

nvidia · hf · 2026-09-03

NVIDIA presents a specialization pipeline combining curated problems, synthetic reasoning, supervised fine-tuning, and reinforcement learning. With iterative test-time refinement, the resulting competitive programming models exceed top human scores on IOI benchmarks, reaching gold-medal-level performance.

Original post →

More from Research

Research channel →