Reasoning in Language Models

Different interaction methods with language models reveal varying world modeling capabilities. While inference time allows models to compute implications of facts, training time often leads to mere memorization. By distilling user preferences into general principles, there's potential to enhance model training and reasoning, paving the way for richer feedback mechanisms.