Deep Papers

Reinforcement Learning in the Era of LLMs

Mar 15, 2024
Exploring reinforcement learning in the era of LLMs, the podcast discusses the significance of RLHF techniques in improving LLM responses. Topics include LM alignment, online vs offline RL, credit assignment, prompting strategies, data embeddings, and mapping RL principles to language models.
Ask episode
Chapters
Transcript
Episode notes