4 ms·
I see the quantized model is supplied. > Note: the full model on GPU (16GB of RAM required) performs much better in our qualitative evaluations. Is there a do
by jakecopp 4y ago
I see the quantized model is supplied.
> Note: the full model on GPU (16GB of RAM required) performs much better in our qualitative evaluations.
Is there a download for a trained full model?
- meghan_rain 4y agojust dequantize it
- akrymski 4y agoArguably the funniest comment I've seen in a while
- stu2b50 4y agoThey have the LoRA delta weights on huggingface, which is linked on the github. Since it's the just the deltas, they're substantially smaller (~8mb), and you'll need to supply the original 7b LLaMA yourself.