Published Sep 17, 2024
Reverse engineering OpenAI's o1
Nathan Lambert dives into OpenAI's o1 model, discussing its future implications, the reinforcement learning principles it utilizes, and the innovative generative strategies it employs to shape the AI landscape.

Topics covered
Popular Clips
Episode Highlights
Related Episodes

OpenAI's Model (behavior) Spec, RLHF transparency, and personalization questions
Answers 383 questions
OpenAI chases Her
Answers 383 questions
AI for the rest of us
Answers 383 questions
AGI is what you want it to be
Answers 383 questions
Stop "reinventing" everything to "solve" alignment
Answers 383 questions
A realistic path to robotic foundation models
Answers 383 questions
Why we disagree on what open-source AI should be
Answers 383 questions
OLMoE and the hidden simplicity in training better foundation models
Answers 383 questions
Text-to-video AI is already abundant
Answers 383 questions
Open Language Models (OLMos) and the LLM landscape
Answers 383 questions
Alignment-as-a-Service: Scale AI vs. the new guys
Answers 383 questions
How to cultivate a high-signal AI feed
Answers 383 questions
