John discusses the concept of fooling images in deep learning models, revealing how easily classifiers can be deceived with imperceptible noise. The conversation highlights the need for formal guarantees in model correctness and sheds light on the limitations of current learning techniques.