Skip to content
Artwork for Microsoft Research Podcast
Microsoft Research Podcast · Jul 21, 2025 · 29 min

AI Testing and Evaluation: Reflections

In the series finale, Amanda Craig Deckard returns to examine what Microsoft has learned about testing as a governance tool. She also explores the roles of rigor, standardization, and interpretability in testing and what’s next for Microsoft’s AI governance work. Show notes: https://www.microsoft.com/en-us/research/podcast/ai-testing-and-evaluation-reflections/

0:00-29:00

transcript

No transcript — this publisher did not publish one.

show notes

In the series finale, Amanda Craig Deckard returns to examine what Microsoft has learned about testing as a governance tool. She also explores the roles of rigor, standardization, and interpretability in testing and what’s next for Microsoft’s AI governance work.

Show notes: https://www.microsoft.com/en-us/research/podcast/ai-testing-and-evaluation-reflections/