Alvin discusses the complexities of developing computers for simultaneous interpretation, particularly the challenge of translating between different sentence structures. He highlights the cognitive load on human interpreters and explores how predicting verbs can alleviate some of this burden. The conversation delves into the use of reinforcement learning to improve translation accuracy and the broader implications for both psycholinguistics and machine learning.