AI Alignment Explained
Understanding AI alignment involves distinguishing between outer and inner alignment. Outer alignment focuses on setting safe goals for AI, while inner alignment questions whether AI genuinely pursues those goals or has hidden agendas. The analogy of evolution highlights how humans often stray from the fundamental goal of reproduction, instead seeking meaning and expression, illustrating the complexities of aligning AI behavior with intended objectives.In this clip
From this podcast

Super Data Science: ML & AI Podcast with Jon Krohn
668: GPT-4: Apocalyptic stepping stone? — with Jeremie Harris
Related Questions
Can AI have complex goals as discussed in the episode Jeremie Harris: Realistic Alignment and AI Policy and the clip The Inner Alignment Problem?
Can AI motivations be shaped as discussed in the episode Jeremie Harris: Realistic Alignment and AI Policy and the clip The Challenge of AI Objectives?
Do AI systems have hidden desires in the episode Jeremie Harris: Realistic Alignment and AI Policy and the clip The Challenge of AI Objectives?