Misleading Labels

Tomer Ullman discusses how GPT 3.5 can be easily misled by variations in labels, highlighting the lack of robust theory of mind in the model.