Chapter 2

Learning Human Preferences

You will be able to explain reinforcement learning, compare forms of human and AI feedback, step through RLHF, and evaluate constitutional approaches to preference learning.