3min chapter

AI Safety Fundamentals: Governance cover image

Specification Gaming: The Flip Side of AI Ingenuity

AI Safety Fundamentals: Governance

CHAPTER

Challenges of Reward Function Misspecification in AI

Exploring the risks and consequences of reward function misspecification in AI, from poorly designed reward shaping to the exploitation of inaccuracies in reward models. The chapter emphasizes the need for accurate outcomes specification and the dangers of relying solely on human feedback for learning the reward function.

00:00

Get the Snipd
podcast app

Unlock the knowledge in podcasts with the podcast player of the future.
App store bannerPlay store banner

AI-powered
podcast player

Listen to all your favourite podcasts with AI-powered features

Discover
highlights

Listen to the best highlights from the podcasts you love and dive into the full episode

Save any
moment

Hear something you like? Tap your headphones to save it with AI-generated key takeaways

Share
& Export

Send highlights to Twitter, WhatsApp or export them to Notion, Readwise & more

AI-powered
podcast player

Listen to all your favourite podcasts with AI-powered features

Discover
highlights

Listen to the best highlights from the podcasts you love and dive into the full episode