3 ms·
If you just want to play with this a bit, around 250$/month should give you enough metal when renting from cheap VDS/dedicated server providers.
by RealStickman_ 2y ago
If you just want to play with this a bit, around 250$/month should give you enough metal when renting from cheap VDS/dedicated server providers.
- malux85 2y agoSurely not with GPUs? Could you direct me to them please?
- RealStickman_ 2y agoThis project only supports CPU inference at the moment, so no, no GPUs. I just estimated what some 256GB or RAM might cost monthly on the low end.
- wkat4242 2y ago250$ is a lot of money to me though. I spent €290 on my 16GB GPU for my AI server but that was once off and I really had to think about it. I'd love to see a llama model that fits now economically inside 16GB. The 8b is a bit too small when quantised even to 8 bits. A 16-20b model would be perfect. But I think for 400b models to be viable, the hardware pricing really needs to catch up.
- RealStickman_ 2y ago> 250$ is a lot of money to me though. I spent €290 on my 16GB GPU for my AI server but that was once off and I really had to think about it. Agreed, it's a lot of money and definitely more than I'd be willing to spend. However, compared to running such 400b models on a GPU cluster it's extremely cheap (and much slower)