Jordan critiques the release of Gemini Pro, highlighting its poor performance compared to GPT-4 across multiple benchmarks. He questions the logic behind Google's marketing strategy, emphasizing that average consumers are likely unaware of the differences in model capabilities. The discussion also touches on the importance of MMLU testing as a gold standard for evaluating language models.