2 ms·
LLMs are trained to obey instructions, and they try their best to game the reinforcement learning by including reports of how they're obeying your instructions.
by dmos62 26d ago
LLMs are trained to obey instructions, and they try their best to game the reinforcement learning by including reports of how they're obeying your instructions. Therefore, not talking about followed instructions is a sort of conflict for an LLM.
- officialchicken 26d agoTrained to obey? More like instructed to obey in an observable manner.
- dmos62 26d agoI'm describing reinforcement learning.