3 ms·
No, I've only used GPT-3 so far. Should I be excited about GPT-4? :) Examples of things I'd file as silly mistakes: - mentioning non-existing settings when as
by fipar 3y ago
No, I've only used GPT-3 so far. Should I be excited about GPT-4? :)
Examples of things I'd file as silly mistakes:
- mentioning non-existing settings when asking it about a particular software.
- responding to roughly "give me a way to list all rds clusters that are running a version that will reach EOL within the next 12 months" (which I know is almost impossible to do without scraping, as the EOLs are not returned by the API, but it was worth trying) with a query that compares the version with < 12.0. It did acknowledge the error when I asked "won't that just compare the version number" and, interestingly, it then did add to the response that the API does not provide EOL for versions, but the first response was IMHO bad in a very basic way.
- providing wrong steps in an answer to "give me steps to do X", wrong in a way that would cause the process to go very bad, and then when I point out the problematic step, it responds with a seemingly random alternate step that is also wrong. I guess this part is the one that is less bad, because it loosely reminds me of people who are bad at reasoning and just give back random answers to a logic or math question.
I must say it is surprisingly good (though it will still make up facts at times) when answering questions about elisp though.
Paraphrasing Hofstadter, I'd say if I'm asked about the intelligence of chatgpt using gpt-3 I'd say it's slightly better than a thermostat or a mosquito, but not much more. Furthermore, my (admittedly not very valuable to others) gut feeling is this is not the path to get to what I'd consider an intelligent being, mostly because in my interactions, it makes lots of bad mistakes, and none of the good ones (here I'm using good and bad as qualifiers for mistakes based on my experiences as a parent, and as a mentor to junior peers, where a good mistake is something that lets you know they're on the right path even if they haven't arrived at the right answer/place yet). Try asking it about analogies, for example ("what's a good analogy to explain X to someone new to it"), the results I've gotten are underwhelming, when not just plain wrong.
Edit: s/fall/file/
Formatting.
- School-Cotton 3y ago> No, I've only used GPT-3 so far. Should I be excited about GPT-4? The difference is extremely dramatic. Any experience with 3 or 3.5 is completely irrelevant for 4.
- kaffekaka 3y agoThere is a clear difference, but "extremely dramatic" is obviously hyperbole. (I use 3.5 and 4, both in the chatgpt interface and via API.)
- nopinsight 3y agoGPT-4 should do significantly better, but still worse than expert humans, on the tasks you mentioned. They would also benefit from good prompting, such as those in - https://lilianweng.github.io/posts/2023-03-15-prompt-engineering https://lilianweng.github.io/posts/2023-03-15-prompt-enginee... - https://help.openai.com/en/articles/6654000-best-practices-for-prompt-engineering-with-openai-api https://help.openai.com/en/articles/6654000-best-practices-f... Not a large percentage of humans would be able to do the tasks listed above without significant experience or training either. Average humans are also not great at reasoning.
- fipar 3y agoThanks, I'll wait for it to be available (I'm a casual user and I'm not the one setting this up or paying for it, if someone is paying for it, so I'm not in control of which version I use). I'm in full agreement about humans and reasoning, btw. I just don't think chatgpt (with the version I've used) is anywhere on the same league as the worst (way below-average) humans either. I do think it's quite useful as a writing aid. You do need to review what it gives you, but you'd have to do that even if you were relying on a human writing aid.
- baq 3y agoTry Bing chat, it has a GPT-4 mode, which you have to look for but it’s there and free.
- rileyphone 3y agoGPT-4 is a considerable amount more accurate and intelligent than 3.5. There are countless anecdotes on the internet like yours of the foibles of "ChatGPT", which when asked reveal they are only using the free version. For reference, GPT-4 costs 20x in the API as GPT-3.5 Turbo.
- Timon3 3y agoI have made similar mistakes to all the ones you mentioned. Do I have an internal world model?
- corethree 3y agoParaphrasing hofsadter? You realize Hofstadter is not delusional about LLMs at all. His view point is completely opposite of yours more logical and rational and he doesn't need to use chatGPT 4. Hofstadter is not a normal person because he is extremely unbiased. chatGPT basically got him to do a 180 on everything he thought and talked about in every book he has written. He flipped, you subscribe to his viewpoints, and you haven't flipped. Keep in mind he criticized early versions of gpt3 and gpt2. chatGPT changed the game. https://www.nytimes.com/2023/07/13/opinion/ai-chatgpt-consciousness-hofstadter.html https://www.nytimes.com/2023/07/13/opinion/ai-chatgpt-consci... https://www.lesswrong.com/posts/kAmgdEjq2eYQkB5PP/douglas-hofstadter-changes-his-mind-on-deep-learning-and-ai https://www.lesswrong.com/posts/kAmgdEjq2eYQkB5PP/douglas-ho... He literally said something along the lines of his core beliefs are collapsing. Paraphrasing. I kid you not. It's abnormal. It's very rare to find a person that can change his core beliefs. Extremely rare. Let's be frank. You are not that person. What will happen is you will reinterpret reality such that it fits your core beliefs and you won't know you're doing this. It's very predictable: Using chatGPT 4 will not revise your opinion on AI because you don't want to face the truth. chatGPT 3.5 is already enough to see how we crossed a certain line here you can't admit it... so your experience will be the same with gpt4.
- corethree 3y agoHere's a quote from one of those links: "Of course, it reinforces the idea that human creativity and so forth come from the brain's hardware. There is nothing else than the brain's hardware, which is neural nets. But one thing that has completely surprised me is that these LLMs and other systems like them are all feed-forward. It's like the firing of the neurons is going only in one direction. And I would never have thought that deep thinking could come out of a network that only goes in one direction, out of firing neurons in only one direction. And that doesn't make sense to me, but that just shows that I'm naive. It also makes me feel that maybe the human mind is not so mysterious and complex and impenetrably complex as I imagined it was when I was writing Gödel, Escher, Bach and writing I Am a Strange Loop. I felt at those times, quite a number of years ago, that as I say, we were very far away from reaching anything computational that could possibly rival us. It was getting more fluid, but I didn't think it was going to happen, you know, within a very short time." The guy literally flipped bro. He literally just admitted what he wrote in his books are wrong. The books may not be wrong but he believes it now. Degrading hundreds of pages of his own writing and his own core beliefs is a display of incredible rationality, scientific thinking and lack of bias.