Published Aug 27, 2018

67 - GLUE: A Multi-Task Benchmark and Analysis Platform, with Sam Bowman

Sam Bowman delves into the GLUE benchmark, a pioneering framework for evaluating natural language understanding models, exploring its impact on model architecture, task selection, and the pivotal role of diagnostic datasets in assessing linguistic capabilities.
Episode Highlights
NLP Highlights logo

Popular Clips

Episode Highlights

Related Episodes