Evaluating AI Metrics
Jack discusses the shift from single metric testing to diverse testing suites in AI evaluation, highlighting the need for varied methodologies. Seth and Jack explore the evolving concept of human baselines in AI performance, questioning the current standards and redefining human-level benchmarks.In this clip
From this podcast

Last Week in AI
Measurement in AI Policy: Opportunities and Challenges
Related Questions