3 ms·
That's true for production models, but a lot of research involves working with fine-tuned models made for one experiment. RL involves constantly updating weight
by tlb 6d ago
That's true for production models, but a lot of research involves working with fine-tuned models made for one experiment. RL involves constantly updating weights. Those may well run in the same cluster as the eval.