Evaluating AI Outputs
The discussion highlights the importance of breaking down outputs into measurable components, making evaluation more objective and scalable. Both Elad and Kanjun emphasize the challenges of developing reliable AI agents, noting that while code reasoning is easier to assess, richer evaluation methods are necessary for broader applications. They also touch on the balance between product development and research, acknowledging that timing and technology readiness are crucial for success.In this clip
From this podcast

No Priors
No Priors Ep. 41 | With Imbue Co-Founders Kanjun Qiu and Josh Albrecht
Related Questions