3 ms·
1.8-3.3x faster Embedding finetuning now in Unsloth
- electroglyph 9mo agosee also: https://www.reddit.com/r/LocalLLaMA/comments/1qk9vmv/1833x_faster_embedding_finetuning_now_in_unsloth/ https://www.reddit.com/r/LocalLLaMA/comments/1qk9vmv/1833x_f...
- danielhanchen 9mo agoExcited to have collabed on this! Thanks electroglyph for the contrib!
- storystarling 9mo agoDo the memory savings carry over to inference or is this strictly optimizing the backward pass? I'm running embedding pipelines via Celery and being able to squeeze this into lower VRAM would help the margins quite a bit.