3 ms·
> reinforcement learning from human feedback, which is the same method used in ChatGPT Is this confirmed? I thought it was not so.
by noam_compsci 4y ago
> reinforcement learning from human feedback, which is the same method used in ChatGPT
Is this confirmed? I thought it was not so.