4 ms·
> The only way to control the output of a LLM is to essentially rate certain types of responses as better Which is my point. You have to mess with its internal
by TerrifiedMouse 3y ago
> The only way to control the output of a LLM is to essentially rate certain types of responses as better
Which is my point. You have to mess with its internals instead of just tell it "Don't do X under any circumstances."
- famouswaffles 3y agoFirst of all, no you don't have to. Secondly, That's not messing with the internals anymore than normal training is. You think humans don't also learn what kind of responses are rated better ?
- deleted 3y ago[deleted]