4 ms·
Is it possible to prompt this model with two or more texts for each image and get masks for each? Something like this inputs = processor(images=images, text=["
by mksystem 11mo ago
Is it possible to prompt this model with two or more texts for each image and get masks for each?
Something like this inputs = processor(images=images, text=["cat", "dog"], return_tensors="pt").to(device)?