Skip to content
Oct 9Fri
Oct 8Thu
Oct 7Wed
Oct 6Tue
  1. Microsoft Research17

    What AI gets wrong and what failure teaches us

    Microsoft Research's Jennifer Neville discusses how evaluation pushes AI systems beyond traditional benchmarks and why "surprising failures" emerge when models are tested on real user needs. She offers practical guidance for working with current AI systems and explains why examining data matters when results defy expectations.

Oct 5Mon
Oct 2Fri
Sep 24Thu
Sep 18Fri
Aug 31Mon
Jul 7Tue