Latent Space: The AI Engineer Podcast cover image

Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)

Latent Space: The AI Engineer Podcast

00:00

Generic LLM Judges vs Bespoke Reward Models

Kyle argues frontier labs will dominate generic judge models, while bespoke reward models may help niche, specific tasks.

Play episode from 55:30
Transcript

The AI-powered Podcast Player

Save insights by tapping your headphones, chat with episodes, discover the best highlights - and more!
App store bannerPlay store banner
Get the app