2 ms·
rlhf = reinforcement learning from human feedback (had to look it up)
by froh 3mo ago
rlhf = reinforcement learning from human feedback
(had to look it up)
- visarga 3mo agoI think it's more RLVR (reinforcement learning from verified rewards). The RLHF is just to align models to human preferences, meaning to behave nice.
- redanddead 3mo agoWhat makes you say that
- versteegen 3mo agoMore accurate to say RLHF aligns models to human preferences, most significantly to be helpful.