Center for AI Policy Podcast cover image

#10: Stephen Casper on Technical and Sociotechnical AI Safety Research

Center for AI Policy Podcast

00:00

Navigating AI Safety Challenges

This chapter delves into the complexities of reinforcement learning from human feedback (RLHF) in ensuring AI safety, emphasizing the limitations and potential alternatives for better alignment with human values. It highlights the transition towards socio-technical AI safety research and the importance of integrating societal factors to manage risks effectively. The discussion underscores the necessity for robust evaluations and strong institutions to adapt to the complexities of AI integration in society.

Transcript
Play full episode

The AI-powered Podcast Player

Save insights by tapping your headphones, chat with episodes, discover the best highlights - and more!
App store bannerPlay store banner
Get the app