Skip to content
Oct 6Tue
  1. Microsoft Research17

    What AI gets wrong and what failure teaches us

    Microsoft Research's Jennifer Neville discusses how evaluation pushes AI systems beyond traditional benchmarks and why "surprising failures" emerge when models are tested on real user needs. She offers practical guidance for working with current AI systems and explains why examining data matters when results defy expectations.

Oct 2Fri