4 ms·
Full GPT-6B can run if you have 22gb ram (CPU or GPU depending on where you run it). Also can run an 8 bit quantized version pretty easily. This takes ~6gb RAM
by acapybara 4y ago
Full GPT-6B can run if you have 22gb ram (CPU or GPU depending on where you run it).
Also can run an 8 bit quantized version pretty easily. This takes ~6gb RAM.
The results seem far off from GPT-3 but apparently it can get good results when fine tuned.
Bigger models like OPT 66B can run on cloud machines (or a really big local system)
OPT 175B weights are not open but can be applied for.
175B would require something like 500GB RAM if not quantized. That's a lot, but it's possible to build that locally if you have a couple 10's of thousands of dollars.
Wait a few years and 175B on a GPU will be no problem.
- boppo1 4y agoWhat does 'quantized' mean in this context?
- acapybara 4y agoBasically stuff a 32 bit value into an 8 bit value (and lose precision). Apparently it doesn't affect the results significantly. More info: https://github.com/huggingface/transformers/pull/17901 https://github.com/huggingface/transformers/pull/17901