Published Jan 25, 2025
Emergency Pod: Reinforcement Learning Works! Reflecting on Chinese Models DeepSeek-R1 and Kimi k1.5
Nathan Labenz delves into the strategic and creative implications of Chinese AI models DeepSeek-R1 and Kimi k1.5, highlighting their open-source nature, emergent reinforcement learning achievements, and potential to reshape global AI dynamics and challenge existing geopolitical rivalries.

Topics covered
Popular Clips
Questions from this episode
- Asked by 204 people
- Asked by 143 people
- Asked by 133 people
- Asked by 122 people
- Asked by 109 people
- Asked by 107 people
- Asked by 101 people
- Asked by 95 people
- Asked by 90 people
- Asked by 90 people
- Asked by 87 people
- Asked by 83 people
- Asked by 78 people
- Asked by 68 people
Episode Highlights
Related Episodes


The Tiny Model Revolution with Ronen Eldan and Yuanzhi Li of Microsoft Research
Answers 383 questions

Emergency Pod: Mamba, Memory, and the SSM Moment
Answers 383 questions

Seeing is Believing with MIT’s Ziming Liu
Answers 383 questions

Inside China's AI Ecosystem: A View From Beijing
Answers 383 questions

OpenAI Sora, Google Gemini, and Meta with Zvi Mowshowitz
Answers 383 questions

Emergency Pod: o1 Schemes Against Users, with Alexander Meinke from Apollo Research
Answers 383 questions

GPT4 - AI Unleashed w/ ChinaTalk Podcast
Answers 383 questions

The AI Multimodal Revolution with Junnan Li and Dongxu Li of BLIP & BLIP2
Answers 383 questions

The State of AI, from the 80,000 Hours Podcast
Answers 383 questions

The Case for Cautious AI Optimism, from the Consistently Candid podcast
Answers 383 questions

The AI Reasoning Revolution with Ought's Jungwon Byun and Andreas Stuhlmüller
Answers 383 questions

Google’s Multimodal Med-PaLM with Vivek Natarajan and Tao Tu
Answers 383 questions













