Chapter 2
Learning Human Preferences
You will be able to explain reinforcement learning, compare forms of human and AI feedback, step through RLHF, and evaluate constitutional approaches to preference learning.
Chapter 2
You will be able to explain reinforcement learning, compare forms of human and AI feedback, step through RLHF, and evaluate constitutional approaches to preference learning.