4 ms·
Why not fine tuning? Back then, I pit it against GPT-3 playground versions that were available before ChatGPT, and I don't seem to recall RLHF being that much t
by mathteddybear 3y ago
Why not fine tuning? Back then, I pit it against GPT-3 playground versions that were available before ChatGPT, and I don't seem to recall RLHF being that much touted back then. RLHF seems mentioned along with the development of InstructGPT.
While it is possible that more RLHF would improve it, let's not jump to the conclusion a bit too fast. Considering that you think that Google wouldn't have resources to fund it, a rather ludicrous notion.