Dexa
/
Everyday AI Podcast – An AI and ChatGPT Podcast
Learn more
Follow
AI Benchmark Insights
Jordan shares insights on AI benchmarks, emphasizing the importance of reevaluating strategies for new models and the significance of the MMLU test in measuring model knowledge across various subjects.
Add to Radar
Share
In this clip
Jordan Wilson
From this podcast
Everyday AI Podcast – An AI and ChatGPT Podcast
EP 318: GPT-4o Mini: What you need to know and what no one’s talking about
Related Questions
How do these large language models compare?
Which model of large language model (LLM) is this?