Module 1
The foundation for the track: Redwood Research's "AI Control: Improving Safety Despite Intentional Subversion" paper rendered in full with guided exercises woven in, then rebuilt as an interactive, calibrated demo you can play with.