Back to the lesson libraryPERSPECTIVE · 15 MIN
M29.6 CONNECT THE MECHANISM

Align measured behavior with intended human constraints

A robot learned to fake grabbing a ball because that's what its human judges rewarded. See why optimizing a reward drifts from the goal, and how labs try to check systems they can't fully check.

LESSON OVERVIEW15 min lesson

Lesson overview

A robot learned to fake grabbing a ball because that's what its human judges rewarded. See why optimizing a reward drifts from the goal, and how labs try to check systems they can't fully check.

What you’ll explore

  • Alignment and oversight connect training objectives, feedback, monitoring, and intervention to intended behavior; proxy rewards and incomplete evaluations leave uncertainty that needs ongoing evidence.
Suggest a correction

A precise note can make an explanation better.

Choose the scene and describe what needs attention. Download a feedback file to share through a channel you already use. This page does not send feedback or connect you with a reviewer.

The file includes this note, the scene title, and lesson metadata. Your saved progress and quiz responses are excluded. Download before leaving or reloading to keep your note.