All courses
AI Safety & Alignment
Advanced · 4 lessons · 0 complete
A sufficiently capable AI system that pursues the wrong goal, even one that seemed reasonable on paper, can cause real harm at scale. This course covers what alignment actually means, concrete ways specification and reward can go wrong, why raw capability makes the problem harder rather than easier, and the real techniques labs use today along with their honest limitations.
