Stanford CS329H: Machine Learning from Human Preferences | Autumn 2024 | Ethics
Stanford Online · 71:00
Value alignment is not one technical problem: “getting AI to do what we really want” can mean intentions, revealed preferences, objective best interests, or moral constraints on third parties, and each reading implies...