3 ms·
Except there is a qualitative difference in the class of knowledge that a statistical word generator and an expert system would generate. Just because a LLM _c
by lanternfish 2y ago
Except there is a qualitative difference in the class of knowledge that a statistical word generator and an expert system would generate.
Just because a LLM _can_ offer valuable and insightful information, doesn't mean that it doesn't also hallucinate. The most troubling factor here is that often the hallucinated content also looks like valuable and insightful information, but is just incorrect. This is the use. You have to hold that awareness whenever interacting with these systems.
- outofpaper 2y agoYup exatly. They are dream machines. LLMs without other systems can only work in the flow. The fact that in this word flow the larger LLMs can generate navigation instructions for actual mazes and solve random algorithmicly generated problems doesn't mean they are not hallucinating it just means we're getting wonderfully useful hallucinations.
- refulgentis 2y agoHallucination is a term that means "imagined facts", so it's very hard for me to parse this comment into something meaningful beyond "if we say it always generates hallucination, we can say it always generates hallucinations"
- jrm4 2y agoYou Google a restaurant that appears to be open. You go there, and you find that the restaurant is no longer there. Did Google "hallucinate" a restaurant? Because this is no different.
- refulgentis 2y agoWe can empirically test if hallucination is a good word for communicating this concept, by checking if people describe(d) that as a hallucination (they don't). This is all IMHO, I'm not trying to be difficult or nitpick, I just don't understand the idea as communicated. As applied to LLMs, it sounds like hallucination == could be wrong, and this Google example seems further away even when steel-manning, ex. we don't say all Google results are hallucinated.
- skywhopper 2y agoIt doesn’t mean automatically wrong. It’s just bullshitting. It makes up something that fits a pattern. Depending on the question, the pattern may be right more often than not. If you ask ChatGPT, “hey is the McDonald’s near my house open at 6pm?” It doesn’t know anything about where you are or if there’s a McDonald’s or what its hours are. It will likely hallucinate that sure, it’s open at 6pm. But when it does so is it “right” in a meaningful way?
- jrm4 2y agoYup. IMHO, I think "bullshitting" is a much better word than hallucinating and/or getting it right! Much like real life bullshitters, it is inclined to say something truthful-sounding, but doesn't actually have a strong reliability towards truth per se.
- randomdata 2y agoBullshitting implies intent to deceive. As far as we know, an LLM honestly "believes" (as if you need another rabbit hole) what it says. Delusion, perhaps? Delusion implies a degree of consistency, though. LLMs can be on point one minute and completely off the rails the next even when prompted with the same prompt. Hallucination fits better here as it speaks to the real-time "perception" (there's another one for you). An LLM is not a brain, though, so no matter which analogy you choose, it will come with some flaws. Regardless, "hallucination" has moved past analogy territory and now has its own LLM-specific usage with reasonably wide acceptance so the analogy angle is now moot anyway.
- jhbadger 2y ago>Bullshitting implies intent to deceive Not in the formal sense. The philosopher Harry Frankfurt famously distinguished bullshitting from lying because a liar knows the truth and is trying to hide it where a bullshitter is simply trying to sound convincing and may or may not be telling the truth (and may not even know themselves if they are) https://en.wikipedia.org/wiki/On_Bullshit https://en.wikipedia.org/wiki/On_Bullshit
- deleted 2y ago[deleted]
- airstrike 2y agoThat's not what the GP is arguing, though
- BeetleB 2y ago> Except there is a qualitative difference in the class of knowledge that a statistical word generator and an expert system would generate. There's a lot of difference between the two, and you don't have to treat it as one or another. It's OK to treat it as something in the middle. > The most troubling factor here is that often the hallucinated content also looks like valuable and insightful information, but is just incorrect. This is the use. You have to hold that awareness whenever interacting with these systems. Completely agree, sans the word "troubling". It's not troubling. It is what it is. As long as you keep it in mind when you use it, and treat it as an entity that can be completely wrong, and use it where it's OK to be completely wrong (e.g. when the output is easily verifiable), there's nothing "troubling" with that.