Center for AI Policy Podcast cover image

#10: Stephen Casper on Technical and Sociotechnical AI Safety Research

Center for AI Policy Podcast

00:00

Navigating AI Interpretability and Safety

This chapter explores the evolving field of AI interpretability, emphasizing its importance for ensuring safety and usability in AI systems. It discusses practical techniques and frameworks for interpreting models, highlighting ongoing challenges and advancements. The conversation also touches on the significance of critical evaluation in the face of commercial hype and the complexities introduced by adversarial attacks on AI models.

Transcript
Play full episode

The AI-powered Podcast Player

Save insights by tapping your headphones, chat with episodes, discover the best highlights - and more!
App store bannerPlay store banner
Get the app