AXRP - the AI X-risk Research Podcast cover image

34 - AI Evaluations with Beth Barnes

AXRP - the AI X-risk Research Podcast

00:00

Evaluating AI Capabilities and Threats

This chapter discusses the complexities of assessing AI models in terms of their performance and potential risks. It highlights issues like the lack of a formal definition for 'capability,' the significance of task difficulty, and the challenges of evaluating models that may act autonomously. The conversation also addresses responsible scaling policies and the need for comprehensive frameworks to ensure safety and ethical standards in AI development.

Transcript
Play full episode

The AI-powered Podcast Player

Save insights by tapping your headphones, chat with episodes, discover the best highlights - and more!
App store bannerPlay store banner
Get the app