5 ms·
One day it will though. I think we will also see a lot of weights optimisation/pruning and model architecture optimisations that allow them to run on much less
by tyfon 4y ago
One day it will though.
I think we will also see a lot of weights optimisation/pruning and model architecture optimisations that allow them to run on much less hardware.
- Baeocystin 4y agoLike you say, we can self-host GPT-2 today. And Bloom is a large-scale open source LLM model available for anyone, and has been around for months. Look how rapidly the RAM/hardware requirements for the diffusion generators has dropped over the past couple of months, now that they passed the critical interest threshold. I see no reason not to expect similar here.