Published Apr 8, 2024

Localizing and Editing Knowledge in LLMs with Peter Hase - 679

Peter Hase delves into the challenges of localizing and editing knowledge in large language models, discussing sensitive information management, interpretability using causal tracing, and enhancing AI operations through easy-to-hard generalization and scalable oversight.
Episode Highlights
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) logo

Popular Clips

Episode Highlights

Related Episodes