4 ms·
For every image I try, I get the same response: > This image shows a diverse group of people in various poses, including a man wearing a hat, a woman in a whee
by matja 1y ago
For every image I try, I get the same response:
> This image shows a diverse group of people in various poses, including a man wearing a hat, a woman in a wheelchair, a child with a large head, a man in a suit, and a woman in a hat.
No, none of these things are in the images.
I don't even know how to begin debugging that.
- exe34 1y agoMeans it can't see the actual image. It's not loading for some reason.
- aendruk 1y agoI’m having a hard time imagining how failure to see an image would result in such a misleadingly specific wrong output instead of e.g. “nothing” or “it’s nonsense with no significant visual interpretation”. That sounds awful to work with.
- tough 1y agoFun fact,you can prompt the llm's with no input and random nonsense will come out of them
- exe34 1y agoAnd if you set the temperature to zero, you'll get the same output every time!
- sigmaisaletter 1y agoLLMs have a very hard time saying "I am useless in this situation", because they are explicitly trained to be a helpful assistant. So instead of saying "I can't help you with this picture", the thing hallucinates something. That is the expected behavior by now. Not hard to imagine at all.
- aendruk 1y agoNo controls in the training data?
- clueless 1y agoI get the same as well, instead I get this message, no matter which image I upload: "This is a humorous meme that uses the phrase "one does not get it" in a mocking way. It's a joke about people getting frustrated when they don’t understand the context of a joke or meme." Not sure why it's not working
- clueless 1y agoOk, following the following comment in this thread fixed the issue: https://news.ycombinator.com/item?id=43943624 https://news.ycombinator.com/item?id=43943624