Reasoning in Language Models
Different interaction methods with language models reveal varying world modeling capabilities. While inference time allows models to compute implications of facts, training time often leads to mere memorization. By distilling user preferences into general principles, there's potential to enhance model training and reasoning, paving the way for richer feedback mechanisms.In this clip
From this podcast

The Gradient
Jacob Andreas: Language, Grounding, and World Models
Related Questions