Published Nov 18, 2022

StrategyQA and Big Bench

Explore the cutting-edge world of AI language models with discussions on the Big Bench benchmark's role in identifying model capabilities, the complexities of crowdsourcing unbiased datasets, and the innovative StrategyQA dataset crafted for implicit reasoning challenges, featuring insights from expert Mor Geva.
Episode Highlights
Data Skeptic logo

Popular Clips

Episode Highlights

Related Episodes