Skip to content
Trending storyArchived

Microsoft Research on AI failures

1 reports1 sourcesUpdated 4 days ago

What you need to know

AI summary

Microsoft Research's Jennifer Neville says evaluation should push AI systems beyond traditional benchmarks, and that "surprising failures" emerge when models are tested against real user needs. She offers practical guidance for working with current AI systems and says examining data matters when results defy expectations.

Generated by AI from the reporting · updated 3 days ago

Timeline

Follow the reports to see every side of the story.

Oct 6
  1. Microsoft Research
    What AI gets wrong and what failure teaches us

    Microsoft Research's Jennifer Neville discusses how evaluation pushes AI systems beyond traditional benchmarks and why "surprising failures" emerge when models are tested on real user needs. She offers practical guidance for working with current AI systems and explains why examining data matters when results defy expectations.