logoalt Hacker News

dmos62today at 9:48 AM1 replyview on HN

LLMs are trained to obey instructions, and they try their best to game the reinforcement learning by including reports of how they're obeying your instructions. Therefore, not talking about followed instructions is a sort of conflict for an LLM.


Replies

officialchickentoday at 11:26 AM

Trained to obey? More like instructed to obey in an observable manner.

show 1 reply