Assignment 5

Post-training

How a pretrained model learns from demonstrations, feedback, and rewards.

Stanford assignment materials

← All tracks