15 ms·
It seems like a hack to be honest. Problem at hand is not to make transformers do addition of 100 digit numbers. Problem is the current systems can’t reason abo
by msoad 2y ago
It seems like a hack to be honest. Problem at hand is not to make transformers do addition of 100 digit numbers. Problem is the current systems can’t reason about things, math included.
Optimizing for a certain use case is not gonna take us where we wanna be. We want to have a system that can learn to reason.
- josehackernews 2y agohow do you argue that these models are not able to reason? deductive reasoning is just drawing specific conclusion from general patterns. something I would argue this models can do (of course not always and are still pretty bad in most cases) the point i’m trying to make is that sometimes reasoning is overrated and put on the top of the cognitive ladder, sometimes I have seen it compared to self-awareness or stuff like that. I know that you are not probably saying it in this way, just wanted to let it out. I believe there is fundamental work still to be done, maybe models that are able to draw patterns comparing experience, but this kind of work can be useful as make us reflect in every step of what these models do, and how much the internal representation learned can be optimized
- YeGoblynQueenne 2y ago>> deductive reasoning is just drawing specific conclusion from general patterns. This is according to whom, please?
- nicklecompte 2y agoThe fundamental argument of "Artificial Intelligence, Natural Stupidity" is that AI researchers constantly abuse terms like "reasoning," "deduction," "understanding," and so on, deluding others and themselves that their machine is almost as intelligent as a human when it's clearly dumber than a dog. My cats don't need "general patterns" to form deductions, they deduce many sophisticated things (on their terms) with n=1 data points. In the 80s the computers were indisputably dumber than ants. That's probably not true these days. But the decades-long refusal of most AI researchers to accept humility about the limitations of their knowledge (now they describe multiple-choice science trivia as "graduate level reasoning") suggests to me that none of us will live to see an AI that's smarter than a mouse. There's just too much money and ideology, and too little falsifiability.
- YeGoblynQueenne 2y agoDrew McDermot's warning is well-heeded, but there are established and well-understood definitions of deductive, inductive and abductive reasoning that go back to at least Charles Sanders Pierce (philosopher and pioneer of predicate logic, contemporary of Gotlob Frege) that are widely accepted in AI research, and that even McDermot would have accepted. See sig for intro.
- nicklecompte 2y agoThis is completely irrelevant. McDermot's point was that scientifically-plausible definitions of reasoning were not actually being used in practice by AI researchers when they made claims about their systems. That is just as true today.
- YeGoblynQueenne 2y agoI've read McDermot's paper a few times (it's a favourite of mine) and I don't remember that angle. Can you please clarify why you say that's his point?
- naasking 2y ago> My cats don't need "general patterns" to form deductions, they deduce many sophisticated things (on their terms) with n=1 data points. No they don't. That's just generalization, so they've seen plenty of other data points that are similar enough.
- fishe 2y agoAnts behave in ways that a modern computer still can't imitate. I don't think that generalized intelligence is possible but if it is it would need a different starting point than our current computing hardware. Even insects are flexible in ways that computers aren't.
- foolswisdom 2y ago> Deductive reasoning is the process of drawing valid inferences. An inference is valid if its conclusion follows logically from its premises, meaning that it is impossible for the premises to be true and the conclusion to be false. <https://en.wikipedia.org/wiki/Deductive_reasoning https://en.wikipedia.org/wiki/Deductive_reasoning>
- YeGoblynQueenne 2y agoThat's not the definition used by the comment above.
- msoad 2y ago> how do you argue that these models are not able to reason? I don't make this argument. Benchmarks like CLUTRR[1] show how poorly LLMs do in reasoning. [1] https://github.com/facebookresearch/clutrr https://github.com/facebookresearch/clutrr
- Last5Digits 2y agoThere is a difference between poor reasoning and no reasoning. SOTA LLMs correctly answer a significant number of these questions correctly. The likelihood of doing so without reasoning is astronomically small. Reasoning in general is not a binary or global property. You aren't surprised when high-schoolers don't, after having learned how to draw 2D shapes, immediately go on to draw 200D hypercubes.
- wrsh07 2y agoGranting that, the original point was that they're not excited about this particular paper unless (for example) it improves the networks' general reasoning abilities. The problem was never "my llm can't do addition" - it can write python code! The problem is "my llm can't solve hard problems that require reasoning"
- mdp2021 2y agoIt is not «deductive reasoning»: it is just "reasoning". That is, revising a body of ideas for qualities pertinent to alethic (truthfulness) and understanding (completeness). It is critical thinking, continuous cycles of reprocessing. And this cannot be overrated: it is the core activity.
- Shrezzing 2y ago>deductive reasoning is just drawing specific conclusion from general patterns. something I would argue this models can do That the models can't see a corpus of 1-5 digit addition then generalise that out to n-digit addition is an indicator that their reasoning capacities are very poor and inefficient. Young children take a single textbook & couple of days worth of tuition to achieve generalised understanding of addition. Models train for the equivalent of hundreds of years, across (nearly) the totality of human achievement in mathematics, and struggle with 10-digit addition. This is not suggestive of an underlying capacity to draw conclusions from general patterns.
- throwthrowuknow 2y agoI think the “train for hundreds of years” argument is misleading. It’s based off of parallel compute time and how long it would take to run the same training sequentially on a single GPU. This assumes an equivalence with human thought based on the tokens per second rate of the model which is a bad measurement because it varies depending on hardware and the closest comparison you could draw to what a human brain is doing would be either the act of writing or speaking but we obviously process a lot more information and produce a higher volume of information at a much higher rate than we can speak or write. Imagine if you had to verbally direct each motion of your body, it would take an absurd amount of time to do anything depending on the specificity you had to work with. The work done in this paper is very interesting and your dismissal of “it can’t see a corpus and then generalize to n digits” is not called for. They are training models from scratch in 24 hours per model using only 20 million samples. It’s hard to equate that to an activity a single human could do. It’s as though you had piles of accounting ledgers filled with sums and no other information or knowledge of mathematics, numbers or the world and you discovered how to do addition based on that information alone. There is no textbook or tutor helping them do this either it should be noted. There is a form of generalization if it can derive an algorithm based on a maximum length of 20 digit operands that also works for 120 digits. Is it the same algorithm we use by limiting ourselves to adding two digits at a time? Probably not but it may emulate some of what we are doing.
- OtherShrezzing 2y ago
- HarHarVeryFunny 2y ago> how do you argue that these models are not able to reason? They just don't have the right architecture to support it. An LLM is just a fixed size stack of N transformer layers, and has no working memory other than the temporary activations between layers. There are always exactly N steps of "logic" (embedding transformation) put into each word output. You can use prompts like "think step by step" to try to work around these limitations so that a complex problem can (with good planning by the model) be broken down into M steps of N layers, and the model's own output in early steps acts as pseudo-memory for later steps, but this only gets you so far. It provides a workaround for the fixed N layers and memory, but creates critical dependency on ability to plan and maintain coherency while manipulating long contexts, which are both observed weaknesses of LLMs. Human reasoning/planning isn't a linear process of N steps - in the general case it's more like an iterative/explorative process of what-if prediction/deduction, backtracking etc, requiring working memory and focus on the task. There's a lot more to the architecture of our brain than a stack of layers - a transformer is just not up to the job, nor was built for it.
- math_dandy 2y agoWe have no definition of reasoning that is sufficiently precise to be useful. But we do have a bunch of benchmark tasks/datasets that test what we intuitively understand to be aspects of reasoning. For AI models, "being able to reason" means "performing well on these benchmarks tasks/datasets". Over time, we'll add more benchmarking tasks and datasets that ostensibly test aspects of "reasoning", and people will develop models that succeed on more and more of these simultaneously. And these models will become more and more useful. And people will still argue over whether they are truly "reasoning".
- sshine 2y ago> Problem is the current systems can’t reason about things Sounds like the AGI argument trap: They're not able to reason, but we can't succintly define what it is. I don't come with a reasoning chip. Whatever I call reasoning happens as a byproduct of my neural process. I do think that the combination of a transformer network and calls to customized reasoning chips (systems that search and deduce answers, like Wolfram Alpha or logic/proof systems) may be a short-stop to something that can perform reason and execution of actions better than humans, but is not AGI.
- short_sells_poo 2y agoI suppose it's a question whether what we call "reasoning" is an emergent phenomenon from having enough connections in a graph, or whether it's some other special sauce which we simply don't have in our current models yet. E.g. humans follow a deductive process to answer questions which they haven't encountered yet. Do we gain this ability purely from a denser/larger graph of knowledge, or from a completely different architecture? I think until we know the answer to this, we can't make predictions about how to build true AGI.
- psychoslave 2y ago> E.g. humans follow a deductive process to answer questions which they haven't encountered yet. Rarely, actually. More generally humans use all kind of inferences where problem at hand is intertwined with all other attention points that is occupying the mental load of the person. Giving a topic full mental attention and finding a path through pure deduction about a circumscribed subject is a rarity, even if you consider only those situations that require any conscious attention at all to perform some action before moving on.
- tadala 2y agoNot within mathematics, where it is the entire sport, and which is the point of contention.
- 2y ago
- baq 2y agoWe're as humanity building a reasoning machine bottom up. It can't reason... yet. Expecting a magical switch that will make it reason about anything and everything is unreasonable. Starting with arithmetic makes perfect sense.
- golol 2y agoAs I understand, conceptually they just changed 346 + 23 = ? to (1: 3, 2: 4, 3: 6) + (1: 2, 2: 3) = ? So it is not that much of a specific hack. There could be a broader principle here where something is holding transformers back in a general fashion, and we might be able to improve on the architecture!
- ckemere 2y agoHopefully 3:3, 2:4, 1:6 and 2:2, 1:3?
- psychoslave 2y agoI didn’t test with all LLM out there, but all of thus I tested failed with something as basic as "What is the number of words in the sentence coming before the next one? Please answer."
- itchyjunk 2y agoHow many humans have you tested this with?
- psychoslave 2y agoInteresting point. Would you please answer the question I was mentioning? :)
- asgeir 2y agoIn my experience, LLMs tend to perform better if you give them instructions before the data to be operated on. At least for the ~13b size models. So,something like: Please count the number of words in the following sentence. "What is the number of words in the sentence coming before the next one?" edit: Which might be an artifact of the training data always being in that kind of format.
- Thorham 2y ago14
- olalonde 2y agoGPT-4 (OpenAI): The sentence you're referring to is "What is the number of words in the sentence coming before the next one? Please answer." It contains 14 words.
- psychoslave 2y agoThanks. I don’t have access to this engine which for some reason is kept in a closed garden for richer people. ¯\_(ツ)_/¯
- 2y ago
- grumpopotamus 2y ago>Problem is the current systems can’t reason about things, math included. Have you tried asking GPT-4 any questions that require reasoning to solve? If so, what did you ask, and what did it get wrong?