LGS comments on How it feels to have your mind hacked by an AI

LGS 29 Mar 2023 22:09 UTC
2 points
0
Not particularly, no. There are two reasons: (1) RLHF already tries to encourage the model to think step-by-step, which is why you often get long-winded multi-step answers to even simple arithmetic questions. (2) Thinking step by step only helps for problems that can be solved via easier intermediate steps. For example, solving “2x+5=5x+2” can be achieved via a sequence of intermediate steps; the model generally cannot solve such questions with a single forward pass, but it can do every intermediate step in a single forward pass each, so “think step by step” helps it a lot. I don’t think this applies to the ice cube question.