Published May 1, 2024

RLHF: A thin line between useful and lobotomized

Nathan Lambert explores the cutting-edge of Reinforcement Learning from Human Feedback, unveiling the potential of innovations like Kahneman-Tversky Optimization to revolutionize AI reasoning and capabilities. With a deep dive into the complexities of preference fine-tuning and chattiness dynamics, Lambert navigates the delicate balance between advanced AI performance and maintaining system integrity.
Episode Highlights
Interconnects Audio logo

Popular Clips

Episode Highlights

Related Episodes