LiveCodeBench
3stories this week
3last 30 days
3all time
Timeline
- Oct 8, 2026 · Research paper · 1 sourceSFT-as-Context Mitigates Forgetting in Supervised Fine-TuningWe introduce SFT-as-context, a training-free method in which the parent model uses the SFT model's response as context to answer the query.
- Oct 7, 2026 · Opinion / analysis · 1 sourceOne Model Family, Two Gold-Level Results: Fine-Tuning Nemotron for IOI and IMOStarting from Nemotron 3, our teams used supervised fine-tuning (SFT), reinforcement learning (RL), and feedback-driven inference to create systems that reached gold-medal level at both IMO 2026 and IOI 2026.
- Oct 6, 2026 · Research paper · 1 sourceUP-MOPD: Update Projection in Multi-Teacher On-Policy DistillationTo address this gap, we propose Update Projection for Multi-Teacher On-Policy Distillation (UP-MOPD).