3 ms·
Yes 2 + 2 is always 4 if you're not a language model and know basic arithmetics
by vhantz 2y ago
Yes 2 + 2 is always 4 if you're not a language model and know basic arithmetics
- CamperBob2 2y agoOr if the language model answers the question by writing and running a Python script. Which is exactly what it can do. Never mind that the tendency to give the exact same answer to the same question over time is not the exhibition of reasoning power you seem to think it is. Have you actually asked some people to multiply 10-digit numbers in their heads? Did they always get the same result? No? Well, there goes that argument. We don't do anything that the LLMs don't do at this point, except adjust our weights (poorly) to move short-term context into long-term memory. Once that capability is added to the models -- which will happen soon enough, because why wouldn't it? -- where will the goalposts go next?
- vhantz 2y agoIt's not about giving the same answer to the same question. It's about getting the right answer 100% of the time, in some very specific domains. If you know, understand and are able to use the basic rules of arithmetic, 2 + 2 only has one answer. If you know, understand and are able to use the basic rules of formal logic, the same premises will lead you to the same conclusion. Two trivial cases for any reasoning person . Two cases that illustrate how fundamentally different LLMs text generation is to reasoning. Two cases that illustrate some of the challenges that need to be solved to bring AI models closer the fiction so many on this site are desperately taking them to be. Of course those who don't care about improving those systems also don't care about understanding their limits, which is unsurprisingly the case for a lot of people on this website.
- CamperBob2 2y agoYou've failed to explain -- or to understand -- how the models get the right answer at all. The fact is, when you ask what 2+2 is, or what 2342+33222 is, the current ChatGPT model will give you the correct answer, even if you don't tell it to write code to get it. The first answer can simply be regurgitated. The second one, not so much. Heck, let's throw in a square root for the fun of it: https://i.imgur.com/Q9eHAaI.png https://i.imgur.com/Q9eHAaI.png How'd it do that, if it can't reason? That problem wasn't in its training corpus. Similar ones were, with different numbers, and that was enough. Ask it 100 times, and it will probably get it wrong a certain percentage of the time... just like you would if I asked you to perform the calculation in your head. Notice that the model actually got the LSD slightly wrong in this example. 188.58 would be a better estimate. It even screws up the way we do. That, to me, is almost as interesting as the fact that it can deal with the problem at all. Of course those who don't care about improving those systems also don't care about understanding their limits The people who do care about improving these systems seem to be doing a pretty awesome job. As for the limits, they frankly don't seem to exist. They certainly aren't where you and your predecessors over the past few years have assured us they are.
- perching_aix 2y agoSo you've never had an encounter at a bar, moderately intoxicated, where you regrettably put the server on the spot by not understanding why you are supposed to pay the amount they're telling you you're supposed to pay? Cause I have, and I have doubts it was significantly more complicated math than basic integer addition and subtraction. I also do use a calculator even for basic, low value, integer math, because what do you know, in my perfectness I often had an issue with numbers not ending up as what they were supposed to. There's also the quite accessible and extensive history of human calculators and the extensive error correction strategies they had to employ because they'd keep cocking up calculations. Come on...