Skip to content
VibeFormer
Advanced28 min

AI Control

Designing deployments that remain safe even if the model is misaligned: monitoring, trusted-untrusted decomposition, and control evaluations.

Not yet written

This lesson is on the syllabus but has no text yet

The full curriculum is published up front so you can see the whole route and its dependencies. Lessons are being written in curriculum order.

What it will cover

  • AI control
  • monitoring
  • untrusted model
  • control evaluation