3 ms·
I agree especially with the second argument. But most people who toy with LLMs will probably never make money out of them. Even those who do will often spend a
by uniqueuid 3y ago
I agree especially with the second argument.
But most people who toy with LLMs will probably never make money out of them.
Even those who do will often spend a lot of time getting their bearings during which the GPU sits idle. Then you begin to ramp up your use but by the time, there's a new generation of GPUs out.
That's why my recommendation is to start with something lightweight.
It's also much less frustrating to start working for a few hours on a rented A100 rather than running into OOMs all the time while fine-tuning batch sizes and waiting for the nth highly quantized model to download.