7 ms·
Which YOLO?
by smallerize 11mo ago
Which YOLO?
- Glemkloksdjf 11mo agoAny current one. they are easy to use and you can just benchmark them yourself. I'm using small and medum. Also the code for using it is very short and easy to use. You can also use ChatGPT to generate small exepriments to see what fits your case better
- throwaway314155 11mo agoThere aren’t any YOLO models for captioning and the other models aren’t robust enough to make for good embedding models.
- Glemkloksdjf 11mo agoYou can get labels out of the classifier and bounding box models. They are super fast. Its just an alternative i'm mentioning. I would assume a person knowing a little bit of that domain. Otherwise the first option would be CLIP i assume. llm-vl is just super slow and compute intensive.