AI coding agents generate more code, but not more software
A study of coding practices across hundreds of firms finds that human code review acts as a significant bottleneck for AI coding tools.
A study of coding practices across hundreds of firms finds that human code review acts as a significant bottleneck for AI coding tools.
NVIDIA reports that fine-tuned Nemotron 3 systems reached gold-medal level at both IOI 2026 and IMO 2026.
Why it matters: The post lays out a four-part specialization recipe and the SFT, RL and inference-loop split behind two gold-level competition results.
GitHub released ReviewBench, an open offline benchmark for AI code review agents.
Why it matters: GitHub's own numbers show how an offline code-review benchmark tracked a production A/B test, useful for teams weighing offline signals.