4 ms·
This discussion is so dumb - finetuning a base model costs ~$1 with LORA/QLORA and can yield same performance as gpt-4, but at 1/100 of the cost per token. Wha
by hallqv 3y ago
This discussion is so dumb - finetuning a base model costs ~$1 with LORA/QLORA and can yield same performance as gpt-4, but at 1/100 of the cost per token.
What Bloomberg did for $10M was not finetuning..
- simonw 3y ago"finetuning a base model costs ~$1 with LORA/QLORA and can yield same performance as gpt-4, but at 1/100 of the cost per token" That's a big claim - can you back that up with any examples?
- Implicated 3y agoI had opened a new tab back when this comment was just a few minutes old in hopes that when I came back there was some really great blog post linked with the details on the sorcery.
- hallqv 3y agohttps://arxiv.org/pdf/2402.00841.pdf https://arxiv.org/pdf/2402.00841.pdf