Published Sep 3, 2021
Evan Hubinger on Effective Altruism and AI Safety
AI researcher Evan Hubinger delves into the complexities of AI safety, highlighting the challenges of interpretability, alignment, and optimization biases in AI systems, while emphasizing the need for transparency and rigorous safety measures to mitigate existential risks.

Topics covered
Popular Clips
Episode Highlights
Related Episodes


Connor Leahy on EleutherAI, Replicating GPT-2/GPT-3, AI Risk and Alignment
Answers 383 questions

Scott Aaronson: Against AI Doomerism
Answers 383 questions

Davidad Dalrymple: Towards Provably Safe AI
Answers 383 questions

Laura Weidinger: Ethical Risks, Harms, and Alignment of Large Language Models
Answers 383 questions

Eric Jang: AI is Good For You
Answers 383 questions

Jeremie Harris: Realistic Alignment and AI Policy
Answers 383 questions

Upol Ehsan on Human-Centered Explainable AI and Social Transparency
Answers 383 questions

Daniel Situnayake: AI on the Edge
Answers 383 questions

Seth Lazar: Normative Philosophy of Computing
Answers 383 questions

Ben Green: "Tech for Social Good" Needs to Do More
Answers 383 questions

Sara Hooker: Cohere For AI, the Hardware Lottery, and DL Tradeoffs
Answers 383 questions

Peter Henderson on RL Benchmarking, Climate Impacts of AI, and AI for Law
Answers 383 questions

Eric Jang on Robots Learning at Google and Generalization via Language
Answers 383 questions

Catherine Olsson and Nelson Elhage: Anthropic, Understanding Transformers
Answers 383 questions

Anant Agarwal: AI for Education
Answers 383 questions
