3 ms·
I'm not an expert on this subject, but I have done a bit of gpt3 finetuning through their api: I think it's clear that "fine tuning" with GPT is different from
by drcode 3y ago
I'm not an expert on this subject, but I have done a bit of gpt3 finetuning through their api:
I think it's clear that "fine tuning" with GPT is different from fine tuning something like Llama2, in that it probably isn't adjusting all the weights of the network, only a tiny subfragment of the network- Exactly how OpenAI accomplishes this is properietary.
The tradeoff is that OpenAI fine tuning is less expensive, but it is also less powerful than "real" fine tuning.
- swyx 3y ago> it probably isn't adjusting all the weights of the network, only a tiny subfragment of the network source please? this actually isnt all that clear to me
- drcode 3y agoIt was what I read on forums when I learned about the process. It's possible that I am mistaken.
- deleted 3y ago[deleted]
- lgvld 3y agoI've been taught in many cases you can indeed fine-tune the last (i.e. closest from the output) layer(s) of a network. Of course, it does not give as good results as fine-tuning the whole model, but it is obviously way less expensive in compute. i.e. you actually don't want your model to re-learn _everything_.