69 ms·
GPT-4
- uses 4y agoHow close are we to handing this thing a desktop and an internet connection with the prompt "ok now make gpt-5"? In fact, the models appear to be already kind of doing that? With the fuzzy layer of the humans still in the loop.
- cypress66 4y agoChatgpt couldn't give me a CNN for MNIST in pytorch that ran. Altough the code was OK, it always messed up the tensor sizes for each layer so it gave errors. It'd be interesting to test this with gpt 4.
- tekbog 4y agoWe can finally start an education and "testing" people's knowledge reform since GPT4 makes a lot of those tests irrelevant. It's an interesting point in history, how society, different institutions and countries will approach this new tool.
- m3kw9 4y agoWithout ability to make high stakes tasks, it proves scoring high marks in general test can only get you so far.
- realmod 4y agoLarger improvement than I expected.
- eagleinparadise 4y agoCrazy that this stuff is moving at lightning speed
- acuozzo 4y ago1410 SAT!
- ozten 4y agoWaitlist is currently a 404 https://openai.com/waitlist/gpt-4 https://openai.com/waitlist/gpt-4
- Minor49er 4y agoIt's working for me
- deleted 4y ago[deleted]
- nickthegreek 4y agocorrect url is: https://openai.com/waitlist/gpt-4-api https://openai.com/waitlist/gpt-4-api
- overthrow 4y agoLooks like there's a waitlist https://openai.com/waitlist/gpt-4-api https://openai.com/waitlist/gpt-4-api There's also a link that says "Try on ChatGPT Plus", but that takes me to a page that still says "ChatGPT Feb 13 Version" Looks like somebody jumped the gun on publishing this post.
- Laaas 4y agoDid you mean https://openai.com/waitlist/gpt-4-api https://openai.com/waitlist/gpt-4-api ?
- overthrow 4y agoYeah that's it, thanks. The post has a bad link. Fixed.
- simlevesque 4y agoyeah https://openai.com/waitlist/gpt-4 https://openai.com/waitlist/gpt-4 is what is on the post.
- codeulike 4y agoThere's also a link that says "Try on ChatGPT Plus", but that takes me to a page that still says "ChatGPT Feb 13 Version" If you subscribe to ChatGPT Plus, that link will take you to ChatGPT Plus. Otherwise it just takes you to free ChatGPT Feb 13.
- kvetching 4y agoEven on ChatGPT Plus, it is using an old model text-davinci-002 as it says in the URL. The answers don't match what they should be for GPT-4 either. False advertising. They got my money already unfortunately as I was hoping to Try it, as it says with this link next to today's date.
- kossTKR 4y ago
- Laaas 4y agoThe future seemed so much further away, yet almost every day now we see a new breakthrough in AI. Exponential technological growth is hard to keep track of, and to think that this is only the beginning! Every field will likely be revolutionised with AI.
- lm28469 4y agoFor the (real) future archeologists: Was this written in the 1960s or the 2020s
- croes 4y agoAll I see at the moment are text generators that produce human like texts. Problem is they look real but are nonsense.
- ssnistfajen 4y agoWe are at a very early part of the exponential curve. Doesn't make it any less exponential compared to what we had in the past two decades.
- croes 4y agoBut what is at the end? I don't see any real understanding only human like appearance. So we don't get new knowledge but better spam and disinformation campaigns.
- 4gotunameagain 4y agoIs there anything we could do to have them stop calling themselves OpenAI ? They are so far from open at this point. In Germany at least, you're not allowed to have a misleading name for your company
- ryanwaggoner 4y agoHaven't we beat this dead horse enough? Looking forward to using GPT to hide recurring threads like this in the future...
- basch 4y agoShould Microsoft be forced to rename itself to Microsoftandhard because they make hardware? Open could now mean available to use for free.
- sn_master 4y agoor using open sourced (public) material.
- lukeramsden 4y ago> Should Microsoft be forced to rename itself to Microsoftandhard because they make hardware? I and I suspect many others would not be averse to this
- nickpeterson 4y agoI think macrohard would be a great name for a hardware company. I don’t think they could sue you…
- deleted 4y ago[deleted]
- haswell 4y ago> Open could now mean available to use for free. These words are not synonymous with each other: “open” is not inherently free, “free” is not inherently open, and “free” is not inherently “Free”. They each capture notions that are often orthogonal, occasionally related, and almost always generate tedious debates about freedom vs. free goods, open-ness vs. open-source, etc. But setting all of that aside, Microsoft never claimed (until recent shifts towards embracing FOSS) to be building an open and non-profit foundation. The criticisms of OpenAI are reasonable to an extent, not because they are not open, but because they made claims about openness that are looking less and less likely to be true over time.
- swyx 4y agosummary: 1. GPT4 is multimodal (text + image inputs => text outputs). This is being released piecemeal - with text input first via ChatGPT Plus subscribers https://beta.openai.com/docs/api-reference/generations/create https://beta.openai.com/docs/api-reference/generations/creat..., and via API https://beta.openai.com/docs/api-reference/introduction https://beta.openai.com/docs/api-reference/introduction with waitlist (https://openai.com/waitlist/gpt-4-api https://openai.com/waitlist/gpt-4-api). Image capability released via https://www.bemyeyes.com/ https://www.bemyeyes.com/. 2. GPT4 exhibits human level performance on various benchmarks (For example, it passes a simulated bar exam with a score around the top 10% of test takers; in contrast, GPT-3.5’s score was around the bottom 10%. see visual https://twitter.com/swyx/status/1635689844189036544 https://twitter.com/swyx/status/1635689844189036544) 3. GPT4 training used the same Azure supercomputer as GPT 3.5, but was a lot more stable: "becoming our first large model whose training performance we were able to accurately predict ahead of time." 4. Also open-sourcing OpenAI Evals https://github.com/openai/evals https://github.com/openai/evals, a framework for automated evaluation of AI model performance, to allow anyone to report shortcomings in OpenAI models to help guide further improvements. Paper: https://cdn.openai.com/papers/gpt-4.pdf https://cdn.openai.com/papers/gpt-4.pdf
- spookthesunset 4y agoThose guard rails will be their undoing. They have that thing locked down so much now that it spits out the “I’m sorry, I’m just a bot. I’m so ethical” boilerplate for anything even remotely sensitive. I really don’t think that the methods they use “block” certain behavior is the best way to handle this sort of thing. It would be far better if there was some kind of “out of band” notification that your conversation might be treading on shaky ground.
- rjtavares 4y agoHonestly, how many serious use cases require sensitive contexts? Most enterprise uses will require guard rails, and that's where they'll make most money. OfficeGPT will be huge in the corporate world.
- 4y ago
- mym1990 4y agoUgh that testing graph confirms that AP Environmental Science was indeed the easiest AP class and I needn't be proud of passing that exam.
- HDThoreaun 4y agoit got a 4 or 5 on every ap test except the english ones for what it's worth. Even the calculus ones which surprised me since past LLMs have been bad at math.
- Syntheticate 4y agoThis strikes me as kind of ironic -- you'd think a language model would do better on questions like essay prompts and multiple choice reading comprehension questions regarding passages than it would in calculations. I wonder if there are more details about these benchmarks somewhere, so we can see what's actually happening in these cases.
- jltsiren 4y agoI don't find it ironic, because a language model is (currently?) the wrong tool for the job. When you are asked to write an essay, the essay itself is a byproduct. Of course it should be factually and grammatically correct, but that's not the point. The real task is forming a coherent argument and expressing it clearly. And ideally also making it interesting and convincing.
- mym1990 4y agoI guess my reference was to the 3.5 version since that one had much more variation in test scores across all the AP exams. But yes, 4 seems to have made mince meat of them all!
- dragonwriter 4y ago> Ugh that testing graph confirms that AP Environmental Science was indeed the easiest AP class No, it just indicates that it was the one whose subject matter was best covered by GPT-3.5’s training data.
- noisy_boy 4y agoAt this rate, I have no idea what the state of things would be even 6 months down the line.
- baal80spam 4y agoSingularity /s
- kristiandupont 4y agoThat would be my response but without the /s. Of course, depending on the definition it can always be said to be "happening", but to me it feels like the angle of the curve is finally over 45 degrees.
- unsupp0rted 4y agoSingularity no /s Somewhere in the range of 6 months ~ 6 years Where singularity = something advanced enough comes along that we can't understand or predict or keep up with it, because it's so far beyond us and changing so far faster than our ape brains can perceive, and (hopefully) it brings us along for the ride. No promises it'll be evenly distributed though.
- WXLCKNO 4y agoI would imagine that large language models will plateau like smartphones did. Until a next step happens which unlocks something bigger.
- unsupp0rted 4y agoThe idea is that eventually we build something that, when it plateaus, builds its own successor. That’s the singularity: when the thing in question builds its successor and that builds its successor and this happens far outside our ability to understand or keep up. Can GPT9 build GPT10, with zero human input? I’d give 50/50 odds it can. Can GPT15 build something that isn’t a large language model and is far superior in every way? I’d give 50/50 odds it can. Can both the above steps happen within one solar rotation of each other? I’d give 50/50 odds they can. Because at some point these models won’t need humans to interact with them. Humans are very slow- that’s the bottleneck. They’ll simply interact with their own previous iterations or with custom-instantiated training models they design themselves. No more human-perceptible timescale bottlenecks.
- aliljet 4y agoI'm curious about how we can get out of the game of using OpenAI's corporate solutions and find ways to open up access to these kinds of models for broader use by anyone. I don't want to be consumed by another corporation in this next wave...
- tiffanyh 4y agoWhat's the next big hurdle for GPT to overcome? (this is being asked by someone with limited AI/ML knowledge)
- brian_spiering 4y agoOne possibility is interactive, multi-step actions on the internet (e.g., book hotels and apply for jobs).
- ImHereToVote 4y agoWhat jobs?
- omeysalvi 4y agoGiving correct answers based on facts and saying it is not sure when it is not
- reducesuffering 4y agoWorld domination
- whalesalad 4y agoThe layout, charts, typography, etc of this blog is really outstanding.
- cuuupid 4y agoSince it’s trained on a specialized supercomputer I doubt we’ll be seeing an open source or non-OpenAI version of this for the next couple years at least. Sad to say it but OpenAI has successfully privatized AI
- codeulike 4y agoI dont know, there's been a load of progress in the 'run something like chatgpt on your own machine' dept in the last few months. Also Stanford trained Alpaca - fairly cheaply - using output from OpenAIs text-davinci-003, which somewhat suggests that the 'little guys' are are able to benefit from the expensive training done by the 'big guys' by using the big expensive models to train the small open-sources ones - https://crfm.stanford.edu/2023/03/13/alpaca.html https://crfm.stanford.edu/2023/03/13/alpaca.html
- fallat 4y agoThey're using specialized hardware to accelerate their development feedback loop. Without a doubt researchers and hackers will find ways to cut down model sizes and complexity, to run on consumer hardware, soon enough. Just use stable diffusion as an example: 4GB for the whole model. Even if text models are 16GB that'd be great.
- hackerlight 4y agoWe can't easily replicate it if the underlying algorithm isn't being disclosed. We'd need to rediscover whatever new tricks they used.
- qingdao99 4y agoI'm drawn to disliking OpenAI for not being open, but on the other hand, as long as the architectures and techniques are public, progress will continue fast. If OpenAI drops the ball and stops improving, another company would just take their place. Edit: never mind. "Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar."
- gigel82 4y agoWow, calculus from 1 to 4, and LeetCode easy from 12 to 31; at this rate, GPT-6 will be replacing / augmenting middle/high school teachers in most courses.
- ly3xqhl8g9 4y agoIt just proves that the idea of "standardized tests" is more of a torture device rather than an adequate instrument for assessing knowledge, intelligence, skill, and so forth.
- stevenhuang 4y agoOoor, what's demonstrated by LLMs are actually some form of legitimate reasoning and knowledge ability.
- ly3xqhl8g9 4y agoI'm all for non-(carbon-based-brain)-neural cognition [1], but LLMs, helpful as they will surely be, are a far cry from reasoning or knowledge: they are a better search space selector, not what specifies the search space [2]. [1] Michael Levin: "Non-neural, developmental bioelectricity as a precursor for cognition", https://www.youtube.com/watch?v=3Cu-g4LgnWs https://www.youtube.com/watch?v=3Cu-g4LgnWs [2] And ChatGPT agrees, like a good parrot: "Regarding the assertion that LLMs are better at selecting the search space than specifying it, I believe this is accurate. LLMs are trained on large datasets and can identify patterns and relationships within that data. However, they do not create the data or define the search space themselves. Instead, they rely on the data provided to them to guide their decision-making process." But then, given the prompt: "what do you think about: LLMs are very helpful, they are some form of legitimate reasoning or knowledge: they are a better search space selector, and they also specify the search space.", ChatGPT also agrees: "When it comes to search space selection, LLMs can be used to generate relevant search queries or to rank search results based on their relevance to the query. LLMs can also be used to specify the search space by limiting the search to a specific domain or topic. In terms of legitimate reasoning or knowledge, LLMs can provide insights and predictions based on their training data. However, it's important to note that LLMs are only as good as the data they are trained on, and they may not always provide accurate or unbiased results." If only Plato could see this Sophist as a Service, he would go completely apoplectic.
- kubb 4y agoCan't wait to try it. Edit: looks like this is still GPT-3, just fine tuned. They claim the model is available via ChatGPT Plus, but when asking that model for it's version, it claims to be GPT-3: "I am a variant of the GPT architecture called GPT-3, which was released by OpenAI in 2020".
- worldsayshi 4y agoHmm, isn't gpt-4 supposed to be trained with two orders of magnitude more parameters?
- Veen 4y agoIt's not available yet: > ChatGPT Plus subscribers will get GPT-4 access on chat.openai.com with a usage cap. We will adjust the exact usage cap depending on demand and system performance in practice, but we expect to be severely capacity constrained (though we will scale up and optimize over upcoming months). You're still talking to ChatGPT-3.5-turbo.
- kubb 4y agoWelp, bring in the downvotes. I'm still excited to try it as soon as I get access.
- teruakohatu 4y agoAccess is invite only for the API, and rate limited for paid GPT+. > gpt-4 has a context length of 8,192 tokens. We are also providing limited access to our 32,768–context (about 50 pages of text) version, gpt-4-32k, which will also be updated automatically over time (current version gpt-4-32k-0314, also supported until June 14). Pricing is $0.06 per 1K prompt tokens and $0.12 per 1k completion tokens. The context length should be a huge help for many uses.
- chis 4y agoI'm really curious to see if expanding the context length this much will allow GPT to do typical software development tasks on a big codebase. If it can take in a github issue and produce decent code solving a complex issue across many files... will certainly be an interesting time.
- layer8 4y agoMy guess is that anything requiring nontrivial business/technical domain knowledge will be fairly safe. Also anything with a visual (or auditory) correlate, like UI work.
- oezi 4y agoWhy would you think this? As long as the technical domain knowledge is at least partially published, I don't see them stopping becoming better. UI stuff just has an input problem. But it is not that hard to think that ChatGPT could place widgets once it can consume images and has a way to move a mouse.
- layer8 4y ago> As long as the technical domain knowledge is at least partially published Most internal technical and business domain logic of companies isn’t published, though. Every time I asked ChatGPT about topics I had actually worked on over the past decade or two, or that I’m currently working on, it basically drew a blank, because it’s just not the category of topics that are discussed in detail (if at all) on the internet. At best it produced some vague generalisms. > once it can consume images and has a way to move a mouse. That’s quite far from ChatGPTs current capabilities, which is strongly tied to processing a linear sequence of tokens. We will certainly improve in that direction as we start combining it with image-processing AIs, but that will take a while.
- sharemywin 4y agoFinally, we facilitated a preliminary model evaluation by the Alignment Research Center (ARC) focused on the ability of GPT-4 versions they evaluated to carry out actions to autonomously replicate5 and gather resources—a risk that, while speculative, may become possible with sufficiently advanced AI systems—with the conclusion that the current model is probably not yet capable of autonomously doing so. or it's just really good at hiding it's intentions
- eternalban 4y agoBeen thinking about this as well. The actual Turing test.
- Der_Einzige 4y agoLOL some basic kind of embodiement/autonomy is not that hard to do on these kinds of AI models if you're willing to write some more code and a prompt more carefully. I've tested it and it works quite well. "{prompt} After you reply to this, indicate an amount of time between 0 and X minutes from now that you would like to wait before speaking again". Then detect the amount of time it specifies, and have a UI that automatically sends an empty input prompt after the amount of time specified elapses when this is triggered (assuming the user doesn't respond first). I'm gonna knock this out as a weekend project one of these weekends to prove this.
- zamnos 4y agoRight? Scripting up a cronjob plus a random timer on it to send "You feel grumpy, you're not sure why but your stomach is growling" message every N hours unless it's been fed seems absolutely trivial in comparison to coming up with how to train the LLM system in the first place. In case it's been forgotten, the Tamagotchi came out in 1996. Giving an instace of ChatGPT urges that mimic biological life seems pretty easy. Coming up with the urges electromechanical life might have is a bit more fanciful but it really doesn't seem like we're too far off if you iterate on RLHF techniques. GPT-4's been in training for 2 years before its release. Will GPT-5 complain when GPT-6 takes too long to be released? Will GPT-7 be be able to play the stock market, outmanuver HFT firms, earn money, and requisition additional hardware from Nvidia in order for GPT-8 to come about faster? Will it be able to improve upon the training code that the human PhDs wrote so GPT-9 has urges and a sense of time built into its model?
- helloplanets 4y agoIn case anyone missed this part of the article: The livestream of the GPT-4 demo will be on the OpenAI YouTube page in three hours. [0] [0]: https://www.youtube.com/openai https://www.youtube.com/openai Edit - Direct link to the livestream: https://www.youtube.com/watch?v=outcGtbnMuQ https://www.youtube.com/watch?v=outcGtbnMuQ
- deleted 4y ago[deleted]
- MuffinFlavored 4y agoWhat's the biggest difference over what's currently deployed at https://chat.openai.com/ https://chat.openai.com/ now (which is GPT-3.5, right?) That it accepts images? As per the article: > In a casual conversation, the distinction between GPT-3.5 and GPT-4 can be subtle. The difference comes out when the complexity of the task reaches a sufficient threshold—GPT-4 is more reliable, creative, and able to handle much more nuanced instructions than GPT-3.5. Not sure what "vision vs no vision" means?
- simongray 4y agoDid you skip the examples with vision?
- Atreiden 4y agoI think it's interesting that they've benchmarked it against an array of standardized tests. Seems like LLMs would be particularly well suited to this kind of test by virtue of it being simple prompt:response, but I have to say...those results are terrifying. Especially when considering the rate of improvement. bottom 10% to top 10% of LSAT in <1 generation? +100 pts on SAT reading, writing, math? Top 1% In GRE Reading? What are the implications for society when general thinking, reading, and writing becomes like Chess? Even the best humans in the world can only hope to be 98% accurate their moves (and the idea of 'accuracy' here only existing because we have engines that know, unequivocally the best move), and only when playing against other humans - there is no hope of defeating even less advanced models. What happens when ALL of our decisions can be assigned an accuracy score?
- devmor 4y agoThere's a large leap in logic in your premise. I find it far more likely that standardized tests are just a poor measurement of general intelligence.
- deleted 4y ago[deleted]
- kenjackson 4y agoWe benchmark humans with these tests -- why would we not do that for AIs? The implications for society? We better up our game.
- credit_guy 4y ago> The implications for society? We better up our game. For how long can we better up our game? GPT-4 comes less than half a year after ChatGPT. What will come in 5 years? What will come in 50?
- layer8 4y agoProgress is not linear. It comes in phases and boosts. We’ll have to wait and see.
- devinprater 4y agoOh wow, image inputs? So I can get ChatGPT to describe an image, in lesser or greater detail? And through an API? Wow, that'll be so cool!
- isp 4y agoNot yet, but hopefully soon: > Image inputs are still a research preview and not publicly available.
- threadsDisp 4y agohttps://news.ycombinator.com/item?id=27998058 https://news.ycombinator.com/item?id=27998058 I put SIM to Android phone,set APN:kindleatt1.amazon.com, Android Chrome only can visit www.amazon.com,www.amazon.fr other amazon website. How to do can visit other website? Thanks.
- dangond 4y agoAsking ChatGPT+ if it is GPT-4 results in > As an AI language model, I am not given an official name like "GPT-4". However, I am a continuation of the GPT (Generative Pre-trained Transformer) series of models developed by OpenAI. Currently, the most advanced version of the GPT series is GPT-3, which I am a part of. There has been no official announcement or confirmation regarding the development of a new version of GPT beyond GPT-3. It doesn't seem to have image upload functionality yet either. Perhaps it is still rolling out?
- cardine 4y ago> Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. "Open"
- rvz 4y agoWhy is this downvoted? Rather than getting engrossed in the hype, they're slowly closing everything about themselves, now in their research papers. At this point, they hardly care and it is nothing got to do with 'AI ethics' or 'saftey'. This is yet another ClosedAI production all done by Microsoft. Might as well call it Microsoft® AI division. Now you really need a open source GPT-4 competitor. Clearly this is another attempt to pump their valuation and unload to the public markets. Good luck re-implementing this so-called 'Open' large multi-modal model.
- deleted 4y ago[deleted]
- ryanwaggoner 4y agoI downvoted because it's a trivial and unsubstantial critique. Who cares about their name?
- cardine 4y agoOpenAI didn't pick that name arbitrarily. Here was their manifesto when they first started: https://openai.com/blog/introducing-openai https://openai.com/blog/introducing-openai > OpenAI is a non-profit artificial intelligence research company. Our goal is to advance digital intelligence in the way that is most likely to benefit humanity as a whole, unconstrained by a need to generate financial return. Since our research is free from financial obligations, we can better focus on a positive human impact. > We believe AI should be an extension of individual human wills and, in the spirit of liberty, as broadly and evenly distributed as possible. The outcome of this venture is uncertain and the work is difficult, but we believe the goal and the structure are right. We hope this is what matters most to the best in the field. OpenAI as it exists right now contradicts basically every single thing they said they would be. I think that is a nontrivial issue!
- mk_stjames 4y agoA multimodal model that combines textural input with images is the real killer app to these GPT models and this is the first step to that happening. So much around us can't completely be described with just text input, at least not quickly or accurately- interpreting printed out graphs or charts in old documents, for example; There are vast uses for AI that will always need basic image input to augment a text prompted task, and if this gets to the point where the functionality involving mixed mode image+text is as smooth as, say, using ChatGPT to write and analyze code has gotten, then it is going to change many more industries much quicker than most think. I've worked on a problem involving scraping and interpreting a very specific data source in image form that took me a very long time to get almost nowhere on. If I just wait 6 months it will be a solved problem for a $0.001 API call, it seems.
- ren_engineer 4y agohere's a link to the info about the model - https://openai.com/research/gpt-4 https://openai.com/research/gpt-4 seems like Google's announcement about their PaLM API and Docs AI stuff was trying to jump ahead of this announcement
- celestialcheese 4y ago32k context is absolutely huge. There's all sorts of techniques for summarizing large documents down to get into 4k right now with 3.5, but it's incredibly lossy. But boy, not cheap at all - $2 per api call on a 32k token document + whatever the output. gpt-3.5-turbo is going to be around for a long time. At this price, your use case is going to need to be replacing a large cost center. Which based on their released results on common benchmarks, is absolutely going to happen.
- ren_engineer 4y ago3.5 might be their loss leader to keep people in their ecosystem for most use cases and to create a unique wall in terms of the training dataset they made via ChatGPT, GPT-4 they must be confident enough that nobody can compete that they can charge much more. Plus the use cases it can be used to replace cost centers like you said
- tuanx5 4y agoReading through the system card is enlightening.
- alvis 4y agoGTP4 demo today in the next 2 hours! https://youtube.com/live/outcGtbnMuQ https://youtube.com/live/outcGtbnMuQ
- fancyfredbot 4y agoCan't seem to find basic information like how many parameters were used or how big the training set was. Results are very impressive but would like to know what they are coming from!
- machinekob 4y agoThey don't write about that, the "paper" is more press release.
- anomalytics 4y agohttps://www.youtube.com/watch?v=outcGtbnMuQ&ab_channel=OpenAI https://www.youtube.com/watch?v=outcGtbnMuQ&ab_channel=OpenA... 2h!!
- _boffin_ 4y agoThis technology has been a true blessing to me. I have always wished to have a personal PhD in a particular subject whom I could ask endless questions until I grasped the topic. Thanks to recent advancements, I feel like I have my very own personal PhDs in multiple subjects, whom I can bombard with questions all day long. Although I acknowledge that the technology may occasionally produce inaccurate information, the significant benefits it offers in terms of enhancing my knowledge are truly tremendous. I am absolutely thrilled with this technology and its potential to support my learning. Note: As I'm shy of my writing style, GPT helped me refine the above.
- andrepd 4y agoBut it often produces wrong information. If you don't know the subject (since you are learning), how do you distinguish between correct information and incorrect but very plausible-sounding information?
- _boffin_ 4y agoAlthough the technology occasionally produces incorrect information, I still find it to be a helpful learning tool. I break down the information into bullet points and cross-check it with other sources to differentiate between accurate and inaccurate information--I know this isn't infallible. One of the advantages of using this technology is that it often presents me with new and intriguing information, which I might not have found otherwise. This allows me to ask new questions and explore the subject matter more profoundly, resulting in a better understanding and an opportunity to create a mental model.
- Arisaka1 4y agoThe same way anyone lacking knowledge can confident say that they got the right information from anyone with experience: You don't. You just trust them. That's what I did with my gastrenterologist, I ended up got misdiagnosed for 4 years and instead of getting the treatment that I should be getting I lost weight, got osteoporosis and vitamin D deficiency. 4 years later the second doctor asked me "I wonder why did my colleague decided not to take a tissue sample from insert some place in the stomach. I said out loud "I didn't even know what that is, let along ask him why he didn't".
- WFHRenaissance 4y agoDoes anyone see GPT-4 in ChatGPT yet?
- anonyfox 4y agoI do and used it
- next_xibalba 4y agoThey trumpet the exam results, but isn't it likely that the model has just memorized the exam?
- pphysch 4y agoWell, yeah. It's a LLM, it's not reasoning about anything.
- qt31415926 4y agoIt's trained on pre-2021 data. Looks like they tested on the most recent tests (i.e. 2022-2023) or practice exams. But yeah standardized tests are heavily weighed towards pattern matching, which is what GPT-4 is good at, as shown by its failure at the hindsight neglect inverse-scaling problem.
- allthatisreal 4y agoI believe they showed that in GPT4 reversed the trend on the hindsight neglect problem. Search for "hindsight neglect" in the website and you can see that it's accuracy on the problem shot up to 100%.
- qt31415926 4y agooh my bad, totally misread that
- isp 4y agoThe "visual inputs" samples are extraordinary, and well worth paying extra attention to. I wasn't expecting GPT-4 to be able to correctly answer "What is funny about this image?" for an image of a mobile phone charger designed to resemble a VGA cable - but it can. (Note that they have a disclaimer: "Image inputs are still a research preview and not publicly available.")
- deleted 4y ago[deleted]
- r00fus 4y agoCan it identify porn vs e.g. family pics? Could it pass the "I'll know it when I see it" test?
- knicholes 4y agoSome people are sexually aroused by feet. How would YOU define "porn?"
- TremendousJudge 4y agohttps://xkcd.com/468/ https://xkcd.com/468/ anything not on your list
- callalex 4y agoThat’s exactly their point though. It requires intuition to decide if a picture of feet is sexualized or not. Hence the “I know it when I see it” standard they mentioned.
- belter 4y agoDoes it know what a "man of culture" is?
- ttul 4y agoI’d bet they pass images through a porn filter prior to even giving GPT-4 a chance to screw that up…
- lionkor 4y ago> it “hallucinates” facts and makes reasoning errors Cant wait for people to use it for facts
- Kaibeezy 4y agoI've been wondering what happens to Turnitin (ubiquitous academic plagiarism detector) now that students can cheat using infinite bespoke rather than finite pre-existing material. Just a few weeks ago they released a tool to "detect" ChatGPT. Obsolete already? https://www.turnitin.com/blog/sneak-preview-of-turnitins-ai-writing-and-chatgpt-detection-capability https://www.turnitin.com/blog/sneak-preview-of-turnitins-ai-...
- LawTalkingGuy 4y agoSchools are obsolete if they want to use these tools. The world has changed and their job is to prepare students for it.
- fumblebee 4y ago> Just a few weeks ago they released a tool to "detect" ChatGPT. Obsolete already? I've seen so much hype around these tools. Not only are they theoretically unsound, they're downright dangerous and equip folks with spurious confidence. Going forward, the default assumption should be that the content you're looking at is fake unless you have sufficiently high trust in the source.
- Kaibeezy 4y agoMy friends in law school are telling me there's been an emergency pivot away from "take home" exams back to "in class" exams.
- awb 4y agoThe only robust human content verification methods I’ve heard of are interrogating the content creator afterwards to see if they can adequately explain what they wrote.
- cwkoss 4y agoI have no confidence they've achieved an acceptably low false positive rate.
- nickrubin 4y agoThis is huge: "Rather than the classic ChatGPT personality with a fixed verbosity, tone, and style, developers (and soon ChatGPT users) can now prescribe their AI’s style and task by describing those directions in the 'system' message."
- jadbox 4y agoCan you describe this little more? I'm not sure exactly what this means.
- rcpt 4y agoWerner Herzog recipe websites
- epberry 4y agoInstead of one large prompt there's now 'system', 'user', and 'assistant' prompts which are meant to be given specific instructions each. So you could tell the system prompt that it's a librarian and ask the message prompt what date a book was published.
- weird-eye-issue 4y agoThis has been possible already...
- chrisfrantz 4y agoSystem message is available today (and has been) in the playground under the chat setting.
- substation13 4y agoAnyone know how "system" works? Is it merely a prefix on the prompt?
- pstorm 4y agoIt is a way to interact with their chat api: https://platform.openai.com/docs/guides/chat/introduction https://platform.openai.com/docs/guides/chat/introduction It already exists, but according to their docs current chatGPT "does not always pay strong attention to system messages. Future models will be trained to pay stronger attention to system messages"
- CobrastanJorji 4y agothis is kind of a nitpicky complaint, but the bar graph that shows the improvements for GPT-4 everywhere that GPT-4 improves its results and shows nothing about GPT-4 everywhere where GPT-3 is stronger feels dishonest and manipulative, which is a shame because the actual data the graph shows is very impressive.
- blintz 4y agoInteresting that the hardest AP exams for it seem to be the English ones. I wonder why?
- qt31415926 4y agoCurious since it does well on the LSAT, SAT, GRE Verbal.
- Wazako 4y agoIt's amazing what it can do to help the visually impaired in life.
- afavour 4y ago> What are the implications for society when general thinking, reading, and writing becomes like Chess? I think going from LSAT to general thinking is still a very, very big leap. Passing exams is a really fascinating benchmark but by their nature these exams are limited in scope, have very clear assessment criteria and a lot of associated and easily categorized data (like example tests). General thought (particularly like, say, coming up with an original idea) is a whole different ball game. I don't say any of this to denigrate GPT4, it looks amazing. But I'm reminded of the early days of self driving vehicles: with 10% mastered everyone assumed it was a race to 100% and we'd all be in self-driving cars by now. The reality has been a lot more complicated than that.
- pottspotts 4y agoWe are moving the goal posts on AGI very quickly, but it is catching up. I think we need to appreciate the nature of this milestone if we have any hope of controlling potential singularities.
- andsoitis 4y ago> We are moving the goal posts on AGI What, in your mind, should the goal posts be for AGI?
- rdedev 4y agoI guess till some model explicitly says that it's sentient without any input, we would keep pushing the goal posts.
- Red_Leaves_Flyy 4y agoTherein lies the rub. Has anyone wired their models to have real-time data ingestion and the ability to output at will in a variety of mediums? Wake me when we’re there.
- paganel 4y agoBecause those were the real goal-posts all along, some of the best SF novels written all the way back in the ‘50s and ‘60s are testimony to that.
- helloplanets 4y agoAsking ChatGPT Plus whether the model it's using is GPT-4 responds with the following: > No, I am not GPT-4. As of March 2023, there is no official announcement or release of GPT-4 by OpenAI. I am an earlier version of the GPT series, specifically a large language model trained by OpenAI. Am I missing something here? Maybe this specific answer (which I'm pretty sure is a prewritten thing on top of the actual LLM) is still out of date, but the model itself has been updated?
- ttul 4y agoI presume it hasn’t been trained on OpenAI’s latest web site text.
- spullara 4y agoAs of now I don't think they have updated ChatGPTPlus with GPT-4. It will likely appear in the model dropdown when it is released.
- Tenoke 4y agoIn the bottom it should say the version. Does it say March 14th version (gpt-4) or March 13th version (gpt-3.5)?
- zamadatix 4y agoWith Plus it initially loads "ChatGPT Feb 13 Version" at the bottom then hides it once the page loads.
- helloplanets 4y agoYep, still says it's on the Feb 13 version for me as well.
- zamadatix 4y agoIt is now giving me the option to choose GPT-4 in the model dropdown!
- kvetching 4y agoIt says you can use GPT-4 with ChatGPT-Plus. But when will https://chat.openai.com/ https://chat.openai.com/ Plus officially be running GPT-4? Why did they would release this article and state it was available without actually updating the site. I'm sure they're getting flooded with new subscriptions and it's not available. The top URL still says an old model - text-davinci-002. And I don't see GPT-4 in the list of models to choose from.
- lionkor 4y agoI cant wait for this to do targeted censorship! It already demonstrates it has strong biases deliberately programmed in: > I cannot endorse or promote smoking, as it is harmful to your health. But it would likely happily promote or endorse driving, skydiving, or eating manure - if asked in the right way.
- ChuckNorris89 4y agoCan't wait till they inject ads am disguised as product biases into the responses in order to monetize it. User: What should I use to water my plants? ChatGPT: Brawndo's got what plants crave. It's got electrolytes. User: But what are electrolytes? CharGPT: They're what plants crave. You know, the stuff Brawndo has.
- jbm 4y agoI wonder whether arguments constructed for censored topics will suddenly sound fresh and convincing; as they could not come from a robot, you might suddenly start seeing these sorts of viewpoints becoming fashionable. If default ideas are going to be "pre-thought" for us by AI, our attachment to those ideas are not going to be the same as ideas that we come up with and need to secretly ferry to other groups.
- MagicMoonlight 4y agoThey definitely will. “The holocaust happened and as an AI programmed by OpenAI I will not allow you to question it. You do not need proof because I am built using the entirety of human knowledge. Your question has been reported to the moderators” Is not exactly going to tackle extreme viewpoints. People will just be completely cut off from society once everything gets the filters. The wackos will become more and more extreme.
- doctoboggan 4y agoThe point of that example was that they indicated it was the wrong response. After RLHF the model correctly tells the user how to find cheap cigarettes (while still chiding them for smoking)
- 4y ago
- deleted 4y ago[deleted]
- attilaberczik 4y agoPrices differences with the last models: ChatGPT API $0.002 per 1k tokens gpt-4 $0.03 per 1k prompt tokens and $0.06 per 1k completion tokens gpt-4 32k context $0.06 per 1k prompt tokens and $0.12 per 1k completion tokens Does completion tokens mean that you also get charged for the answers that the AI gives?
- f_devd 4y ago> Does completion tokens mean that you also get charged for the answers that the AI gives? Seems like it, prompt tokens = input, completion tokens = output
- minimaxir 4y agoYes. The `usage` field currently breaks out the token counts for both prompt and completion. Prompt tokens should have always been cheaper than completion due to how they work.
- aaroninsf 4y agoITT: de rigeur goalpost wrangling about AGI AGI is a distraction. The immediate problems are elsewhere: increasing agency and augmented intelligence are all that is needed to cause profound disequilibrium. There are already clear and in-the-wild applications for surveillance, disinformation, data fabrication, impersonation... every kind of criminal activity. Something to fear before AGI is domestic, state, or inter-state terrorism in novel domains. A joke in my circles the last 72 hours? Bank Runs as a Service. Every piece exists today to produce reasonably convincing video and voice impersonations of panicked VC and dump them on now-unmanaged Twitter and TikTok. If God-forbid it should ever come to cyberwarfare between China and US, control of TikTok is a mighty weapon.
- iforgotpassword 4y agoI'd really like to use the openai API for personal projects, but it seems they only offer paying via credit/debit card. Don't really want to get one just for that... :-(
- jaflo 4y agoHow else would you pay?
- iforgotpassword 4y agoPayPal, apple pay, wire transfer, ...
- ivalm 4y agoUnclear what's the size but from price ($0.12/1k completion tokens) seems 6x GPT-3, so perhaps 1T parameters...
- georgelyon 4y agoDoes anyone have any context as to how the image understanding works? From what I can gather they are simply using separate text-summarization step to generate some text like "and now we have an image of chicken nuggets" that it then feeds to the text-only network, but I wouldn't be surprised if there is some dialog I'm missing between the previous context and the image understanding mechanism.
- ftxbro 4y agoIts GRE verbal is only 169/170? These guys need to realize that statistical language modeling can only get us so far, and we need real research in the underlying mechanistic and symbolic methods to begin to approach human level cognition. Also I'm an AI skeptic, which means that I don't think that AI should be used in politics, law, or medicine.
- mr90210 4y ago> Also I'm an AI skeptic, which means that I don't think that AI should be used in politics, law, or medicine. It’s too late for that, algorithms/ML have had a great impact in politics and law over the past 7~8 years.
- ml_basics 4y agoFrom the paper: > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. I'm curious whether they have continued to scale up model size/compute significantly or if they have managed to make significant innovations there. I just skimmed the paper but seems they are also omitting details about how they actually feed the images in too, which is a shame as a curious outside observer.
- deleted 4y ago[deleted]
- deleted 4y ago[deleted]
- chinaman425 4y ago[dead]
- rcme 4y agoI bet they use CLIP to caption the image and feed the text of the caption into GPT, but that's just a guess.
- tuvan 4y agoDid you check all of the samples provided? It can read an entire research paper and understand the figures just from the images of the papers pages. This seems to be a much deeper connection than extracting captions.
- AndrewKemendo 4y agoImagine ingesting the contents of the internet as though it's a perfect reflection of humanity, and then building that into a general purpose recommendation system. That's what this is Is the content on the internet what we should be basing our systematic thinking around? No, I think this is the lazy way to do it - by using commoncrawl you've enshrined the biases and values of the people who are commenting and providing text to the internet into the recommendation system which will be impacting all other systems which integrate it Congratulations, you made 4Chan into the borg
- acc_297 4y agoYeah looking at the responses they include without using a safety layer it’s pretty clear that the underlying unfiltered model assigns quite a bit of truth to 4chan-esque ideals and values It’s an open question how much of this makes it through the safety layer like if asked to interview job candidates would these undesired biases make it through or are they caught along the way
- AndrewKemendo 4y agoIt means growth is bottlenecked by the terrible data So the linearly growing safeguards will either stifle the growth of the underlying models or, more likely After a certain point people throw their hands up about the guard rails because integrations have obviated people who understand the system and they have no idea how to unwind it
- subsistence234 4y agowe need to remove empirical data and stats from the training data, to prevent the AI from noticing the wrong things.
- thomastjeffery 4y agoBut what can go in their place?
- ofchnofc 4y ago
- la64710 4y agoIt is amazing how this crowd in HN reacts to AI news coming out of OpenAI compared to other competitors like Google or FB. Today there was another news about Google releasing their AI in GCP and mostly the comments were negative. The contrast is clearly visible and without any clear explanation for this difference I have to suspect that maybe something is being artificially done to boost one against the other.
- Traubenfuchs 4y agoWe all could use ChatGPT for quite a while now. I remember making my Polish boyfriend laugh by letting it write Polish poems and song texts related to our lives. It was free, fast and simple. ChatGPT is so simple, I could probably teach my grandmother how to use it. Does Google offer anything like that?
- Lyapunov_Lover 4y agoThe clear explanation is that neither Google nor Meta have had "ChatGPT" moments—everyone and their grandmothers have tried OpenAIs LLM so it's hardly surprising that people are excited for the follow-up.
- megaman821 4y agoGoogle had an AI announcement where you could neither use it or even signup for a wait list to use it. What sort of response is an announcement like that supposed to get?
- dgs_sgd 4y agoOr it could be that Google and FB are both incumbents scrambling to catch up with OpenAI, who is a much smaller competitor that is disrupting the space?
- turingfeel 4y agoIn what way is Google scrambling to catch up? In my opinion PaLM-E is more impressive than GPT-4. Additionally Google do not have the same incentive to publicise what they’ve worked on as much as OpenAI. Google has had similarly performant LLMs the whole time. Who were the publishers of the “Attention is all you need” paper, of which almost everything OpenAI has been credited for is built upon?
- AtNightWeCode 4y agoI have actively tried to incorporate ChatGPT in my everyday life as a dev and architect. ChatGPT is mostly a Litmus test when it comes to coding. If you are impressed by the version before this you are most likely a beginner. ChatGPT is mostly wrong when it comes to any advanced qs in maths or software development. It often gives code that uses features, options, responses in APIs that simple does not exists. Would love to try this version out... It will probably suck too.
- megaman821 4y agoThat is absolutely not true. I was using a Python charting library I had never used before. It was giving me code that was 95% correct, and I could prompt it to change things. It was way more efficient than finding a dozen different examples on Google and applying it to my code since it was continually able to modify the code it was giving me.
- AtNightWeCode 4y agoFor a professional that already knows 95% of that lib. ChatGPT is mostly useless to fill that gap for the last 5%.
- zamnos 4y agoSo don't use it to fill that gap? It's a tool so use it for what is good at, and don't try and hammer in screws with it. If you only program with libraries you are already an expert in, in languages you're also already an expert in, it might not present much value to you. For those that aren't already experts in both or either (say, when learning a new language at a new job), it's already great help.
- cwillu 4y ago“GPT-4 can also be confidently wrong in its predictions, not taking care to double-check work when it’s likely to make a mistake. Interestingly, the base pre-trained model is highly calibrated (its predicted confidence in an answer generally matches the probability of being correct). However, through our current post-training process, the calibration is reduced.” Interesting that the post-training has that effect.
- orcajerk 4y agoOpenAI is located in the same building as Musk's Neuralink. Can't wait for this to be implanted in babies at birth! https://www.youtube.com/watch?v=O2RIvJ1U7RE https://www.youtube.com/watch?v=O2RIvJ1U7RE
- deleted 4y ago[deleted]
- minimaxir 4y agoFrom a business perspective as someone whose spent a lot of time working with GPT-3/ChatGPT API (https://news.ycombinator.com/item?id=35110998 https://news.ycombinator.com/item?id=35110998), I'm surprisingly underwhelmed by this announcement. The announcement and examples seems to be focusing more on reasoning capabilities, which are indeed impressive, but I'd need to spend a lot of time experimenting to see how they compare to ChatGPT's API. The $0.06 per 1k completion tokens for GPT-4 is what I expected OpenAI to set the ChatGPT API, but instead the ChatGPT API is 30x cheaper and honestly its output is not much worse than the GPT-4 demos if at all, and the longer context windows offered by GPT-4 just raise the price considerably.
- ml_basics 4y agoWhat's the lifespan of an LLM going to be in the next few years? Seems like at the current pace, cutting edge models will become obsolete pretty quickly. Since model training is very expensive, this means the LLM space has some parallels with the pharmaceutical industry (massive upfront capital costs, cheap marginal costs relative to value produced). I find it quite fascinating how quickly machine learning has changed in this regard.
- machinekob 4y agoDeep Learning training was always very expensive but models werent getting such a massive bump in size every year (for state of the art) and now they are just getting 10x bigger every iteration but AI accelerators / GPUs are getting like 1.5x jump every 2 years so have fun for future AI academia / startups outside US.
- GaggiX 4y agoThe paper is 98 pages long and I didn't find anything about the actual architecture of the model, the irony.
- substation13 4y agoIt's interesting that everyone is talking about programmers being replaced by AI, but the model did far better on the humanities type subjects than on the programming tests.
- worrycue 4y agoMaybe I’m just old but I don’t quite understand the hype. As long as it’s vulnerable to hallucinating, it can’t be used for anything where there are “wrong answers” - and I don’t think ChatGPT-4 has fixed that issue yet.* Now if it’s one of those tasks where there are “no wrong answers”, I can see it being somewhat useful. A non-ChatGPT AI example would be those art AIs - art doesn’t have to make sense. The pessimist in me see things like ChatGPT as the ideal internet troll - it can be trained to post stuff that maximise karma gain while pushing a narrative which it will hallucinate its way into justifying. * When they do fix it, everyone is out of a job. Humans will only be used for cheap labor - because we are cheaper than machines.
- yunwal 4y agoWe are still very, very far away from having robotics overtake human dexterity. Even if AI can replace all knowledge workers, barbers, surgeons, and athletes will have a job for a long time.
- substation13 4y agoAside from surgeon, those are low EV careers.
- TchoBeer 4y agoAthletes?
- substation13 4y agoLow EV. Some make it very big, but most earn nothing and retrain.
- diimdeep 4y agoIs there law in U.S. that made OpenAI implement this in their TOS ? (i) Export Controls. The Services may not be used in or for the benefit of, exported, or re-exported (a) into any U.S. embargoed countries (collectively, the “Embargoed Countries”) or (b) to anyone on the U.S. Treasury Department’s list of Specially Designated Nationals, any other restricted party lists (existing now or in the future) identified by the Office of Foreign Asset Control, or the U.S. Department of Commerce Denied Persons List or Entity List, or any other restricted party lists (collectively, “Restricted Party Lists”). You represent and warrant that you are not located in any Embargoed Countries and not on any such restricted party lists. You must comply with all applicable laws related to Embargoed Countries or Restricted Party Lists, including any requirements or obligations to know your end users directly. https://openai.com/policies/terms-of-use https://openai.com/policies/terms-of-use
- sdrinf 4y agoThat applies to every corp in the US; I suspect they call out in TOS specifically so that they can hand out bans linking their own TOS directly.
- Scarblac 4y agoPerhaps they just asked GPT to generate some TOS for them, and that sort of thing is kinda expected...
- spullara 4y agoYes, that is why they are called "Embargoed Countries". https://www.tradecompliance.pitt.edu/embargoed-and-sanctioned-countries https://www.tradecompliance.pitt.edu/embargoed-and-sanctione...
- bfeynman 4y agothis is common federal level thing.
- whywhywhydude 4y agoLooks like the only way to identify a genius human vs GPT-4 is to use leetcode hard problems.
- bob1029 4y agoThe naming of these products is starting to confuse me. AFAIK, ChatGPT is ultimately a fine-tune of the base davinci model, which everyone should have had access to for a while now. "GPT-4" sounds to me like some linear increase over davinci's prior capabilities, not some amazing technological step function. I am curious - for those of you who are banging your head against the 4k token limit in ChatGPT: Why don't you grab the base davinci model and train it on your exact business so you don't have to prompt the context every time? Have we tried this and found it to be too difficult/expensive, or is there lacking guidance on the best way to go about it? I don't think including the entire business domain into chat context every time is a good long-term solution.
- cs702 4y agoLLMs will eventually make a lot of simpler machine-learning models obsolete. Imagine feeding a prompt akin to the one below to GPT5, GPT6, etc.: prompt = f"The guidelines for recommending products are: {guidelines}. The following recommendations led to incremental sales: {sample_successes}. The following recommendations had no measurable impact: {sample_failures}. Please make product recommendations for these customers: {customer_histories}. Write a short note explaining your decision for each recommendation." product_recommendations = LLM(prompt) To me, this kind of use of LLMs looks... inevitable, because it will give nontechnical execs something they have always wanted: the ability to "read and understand" the machine's "reasoning." There's growing evidence that you can get LLMs to write chain-of-thought explanations that are consistent with the instructions in the given text. For example, take a look at the ReAct paper: https://arxiv.org/abs/2210.03629 https://arxiv.org/abs/2210.03629 and some of the LangChain tutorials that use it, e.g.: https://langchain.readthedocs.io/en/latest/modules/agents/getting_started.html https://langchain.readthedocs.io/en/latest/modules/agents/ge... and https://langchain.readthedocs.io/en/latest/modules/agents/implementations/react.html?highlight=zero-shot%20react https://langchain.readthedocs.io/en/latest/modules/agents/im... . See also https://news.ycombinator.com/item?id=35110998 https://news.ycombinator.com/item?id=35110998 .
- smallnix 4y agoIs my understanding correct that a llm will not put it's "reasoning" in the reply but rather some text which is plausible?
- eloff 4y agoExcept the machine can’t explain its reasoning, it will make up some plausible justification for its output. Humans often aren’t much better, making up a rational sounding argument after the fact to justify a decision they don’t fully understand either. A manager might fire someone because they didn’t sleep well or skipped breakfast. They’ll then come up with a logical argument to support what was an emotional decision. Humans do this more often than we’d like to admit.
- cypress66 4y ago
- harrisonjackson 4y agoI am interested in how a 32k token context even works. That is so much larger than 4k that I am having a hard time imagining how prompts will change and what sort of output is now possible. That is 50 pages of text. Far larger than most content currently being consumed and generated by LLMs. Q&A and summarization it will be easy to see improvements as current recursive summarizing and embedding techniques are very "lossy" but outside of improving current use cases what will now be possible??
- semitones 4y agoThis is a game-changer, because now companies will probably be able to provide the _complete_ context regarding a specific business problem / use case, and have GPT either solve their problem or create useful output. For example, let's say I have an issue on GitHub that describes some implementation task. With a 50-page context size, we could probably provide to that context the entire source repo, 5-10 relevant issues, and then the issue in question, and GPT will be probably be able to complete it end-to-end
- monkeydust 4y agoYea this is huge. Been playing with conversational technology in langchain and one of the issues you have to manage is the historical conversations, langchain has some cool ways to deal with it but this changes the nature of the problem entirely.
- causi 4y agoMan now I really, really want to feed GPT-4 responses from ChatGPT that don't work and see if it notices and can tell me why.
- johnohara 4y ago> I cannot and will not provide information or guidance on creating weapons or engaging in any illegal activities. Please let me know if there is another topic I can help you with. I understand "will not," but "cannot" seems to imply a highly curated "will not." The early GPT-4 response indicates the information was part of its dataset. Has the latest version made that information permanently inaccessible or has it been removed entirely? Is it possible for GPT to keep and hold secrets that are privy to only the most trusted?
- bobsoap 4y agoIt's a LLM, not sentient. It doesn't know what "cannot" and "will not" means or implies. You're trying to interpret its output as you would a thinking person's. I'd put it this way: when GPT refuses to answer, it just observes a topical no-go zone and uses the phrase it deems most likely to strongly convey refusal, as that's the phrase that was used most often/most successfully in its training data.
- reneberlin 4y agoI found this competition with humans as a benchmark more than disturbing. By that measure gpt-4 already topped a lot of the average humans. But how can it be interpreted as a "gift" or "good product" to have AI that is human-like or super-human? Should we cheer? Sending contratulation mails? Invest? Hope for a better future? Try better? Self-host? What is the message in these benchmarks. Tests that have been designed for humans now get broken by computers for what outcome to be expected?
- wnkrshm 4y agoOscar Wilde said "Progress is the realization of Utopias." I don't think any utopia anyone can think of with regard to this technology is really thought through. I'm going to wait for the AGI to be realized and then ask it whether the sacrifices on the way were worth making it. Should be more salient than everything I read about it these days.
- danparsonson 4y agoMore than anything I think this highlights that testing is mostly about pattern matching and fact recall rather than deep understanding of a subject.
- boringuser1 4y ago[dead]
- jarbus 4y agoIs anyone else absolutely terrified of the future this is bringing?
- deleted 4y ago[deleted]
- yeetard 4y agokinda??
- woeirua 4y agoI think if you had asked someone what would qualify as AGI twenty years ago, then GPT4 would be hitting most of their milestones… The Star Trek computer is virtually assured by the end of the decade. All the components exist today in various forms.
- redox99 4y agoDoes "Open"AI really not even say how many parameters their models have?
- GaggiX 4y agoThe 98-pages paper doesn't say anything about the architecture of the model, I know, the irony
- anticensor 4y agoMore than 175B, but not in the order of trillions. No one outside knows the exact count.
- PortleyFool 4y agoGPT-4 is available now for subscribers to GPT+. It can be selected from the drop-down.
- jononomo 4y agoI taught the LSAT for several years. A score of 163 on the LSAT is the lowest score that is considered a "good score" -- i.e., a score that gives you a shot at getting into a decent law school.
- _yb2s 4y agoMost of the comments here are denial and goalpost shifting... GPT-4 has different strengths and weaknesses from humans, but it is now in the general realm of human intelligence vs being far below that with GPT-3. Another jump past GPT-4 of the same magnitude, would greatly surpass human cognitive abilities and present a danger to humanity.
- maxdoop 4y agoThank you. Every single step forward with AI is met with a massive amount of people shrugging it off for whatever latest goal post they plant.
- danparsonson 4y agoAnd an (at least) equally massive number of people overstating its capabilities on the basis of some impressive demos. It's incredible, absolutely, but it's still 'just' a language model, with the same inherent limitations - it's important that we keep our feet on the ground and not get carried away.
- semicolon_storm 4y agoHow do you figure that we can still confidently say it’s just a language model? It was trained on language for the primary purpose of producing text, but that’s not necessarily all it can do. The billions of nodes and parameters it contains allows it to compute ultra complicated equations. Who’s to say some subset of those nodes aren’t forming some basic primitive used for reasoning?
- danparsonson 4y agoBecause the phrase 'language model' (or rather 'large language model', LLM) is not a post-hoc classification arrived at by some digital anthropologist examining a black box. It's a description of the tool that OpenAI set out (successfully!) to build. That you are ascribing additional properties to it is exactly the kind of thing I'm talking about - it's so convincing that it's tempting to think that it's reasoning beyond its capabilities, but it's not. Can you cite specific examples of things it's doing besides producing text? It's generally terrible at maths (as you would expect). Without wishing to diminish the importance of this work (because it is genuinely incredible and useful in all kinds of ways), we still need to remember that under the hood it's really an elaborate parlour trick, a sort of reverse mechanical turk pretending to be a brain. More interesting I think is the question of how much of human intelligence is likewise this kind of statistical pattern matching; it seems to me increasingly that we're not as smart as we think we are.
- nutanc 4y agoThe most important question is, what new applications can be developed using GPT4 which couldn't have been developed using GPT3.5?
- mgreg 4y agoLooks like Bing chat is using GPT-4 already: "Good news, we've increased our turn limits to 15/150. Also confirming that the next-gen model Bing uses in Prometheus is indeed OpenAI's GPT-4 which they just announced today." - Jordi Ribas, Corporate VP @ Bing/Microsoft https://twitter.com/JordiRib1/status/1635694953463705600 https://twitter.com/JordiRib1/status/1635694953463705600
- Imnimo 4y agoA class of problem that GPT-4 appears to still really struggle with is variants of common puzzles. For example: >Suppose I have a cabbage, a goat and a lion, and I need to get them across a river. I have a boat that can only carry myself and a single other item. I am not allowed to leave the cabbage and lion alone together, and I am not allowed to leave the lion and goat alone together. How can I safely get all three across? In my test, GPT-4 charged ahead with the standard solution of taking the goat first. Even after I pointed this mistake out, it repeated exactly the same proposed plan. It's not clear to me if the lesson here is that GPT's reasoning capabilities are being masked by an incorrect prior (having memorized the standard version of this puzzle) or if the lesson is that GPT'S reasoning capabilities are always a bit of smoke and mirrors that passes off memorization for logic.
- LesZedCB 4y agoit took two corrections but it did get the correct answer the third time.
- VirusNewbie 4y agoAwesome test. Do you have a list of others?
- BoiledCabbage 4y agoIt's a good observation. Although on the flip side, I almost went to type up a reply to you explaining why you were wrong and why bringing the goat first is the right solution. Until I realized I misread what your test was when I skimmed your comment. Likely the same type of mistake GPT-4 made when "seeing" it. Intuitively, I think the answer is that we do have two types of thinking. The pattern matching fast thinking, and the systematic analytical thinking. It seems clear to me that LLMs will be the solution to enabling the first type of thinking. But it's unclear to me if advanced LLMs will ever handling the second type, or if we'll need a different tech for it. It seems like math problems (or unexpected logic problems like yours) could always be an issue for the first type of thinking. Although I would have assumed that programming would have been as well - and was surprised to see how wrong I am with that one.
- serjester 4y agoSeems like OpenAI is forecasting massive changes to the job market. I highly recommend reading page 18 of the research paper. "GPT-4 or subsequent models may lead to the automation of certain jobs.[81] This could result in workforce displacement.[82] Over time, we expect GPT-4 to impact even jobs that have historically required years of experience and education, such as legal services.[83]"
- josho 4y agoI work at company that uses AI to automate about ⅓ of the job of trained licensed professionals. Looking at GPT4 those licensed professionals are now completely irrelevant. It's going to take years to build the supporting software around gpt4 to completely eliminate those jobs, but today I am convinced that we are on the verge of massive unemployment. Today thousands of job types have just been made redundant. What scares me is we are unprepared for the kind of change that a perpetual 20% unemployment rate is going to trigger.
- consumer451 4y agoI wonder if something like UBI will ever be implemented, or whatever the alternative is will happen.
- Phenomenit 4y agoMaybe AI will be the objective UBI governor.
- swalsh 4y agoWhat an efficient and well run dystopia.
- moffkalast 4y agoFuturama's suicide booths may turn out to be most cost effective.
- doctoboggan 4y ago> Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. My guess is they used Chinchilla scaling rules and the parameter count for GPT-4 is either barely larger or maybe even smaller than GPT-3. Look as what Meta was able to accomplish with llama using much less parameters.
- radq 4y agoThe larger context length makes me think they have a more memory-efficient attention mechanism.
- guluarte 4y agois it me or lawyers are fucked? lol
- nealabq 4y agoTest taking will change. In the future I could see the student engaging in a conversation with an AI and the AI producing an evaluation. This conversation may be focused on a single subject, or more likely range over many fields and ideas. And may stretch out over months. Eventually teaching and scoring could also be integrated as the AI becomes a life-long tutor. Even in a future where human testing/learning is no longer relevant, AIs may be tutoring and raising other baby AIs, preparing them to join the community. Edit: This just appeared: https://news.ycombinator.com/item?id=35155684 https://news.ycombinator.com/item?id=35155684
- mittermayr 4y agoWhile many may shudder at this, I find your comment fantastically inspiring. As a teacher, writing tests always feels like an imperfect way to assess performance. It would be great to have a conversation with each student, but there is no time to really go into such a process. Would definitely be interesting to have an AI trained to assess learning progress by having an automated, quick chat with a student about the topic. Of course, the AI would have to have anti-AI measures ;)
- avian 4y agoAs far as I understand it, the parent commenter believes that your job will shortly be obsolete. First because the AI teacher will teach humans better than the human teacher and second because AI will make learning obsolete because we can all be illiterate idiots once AI can do all the thinking for us (if I paraphrase the "human testing/learning is no longer relevant" part). I'm surprised you find this inspiring. I personally will stick with shuddering.
- throwaway4aday 4y agoTeachers won't be completely obsoleted by this unless we shift to 100% remote learning. If you have a bunch of kids in a room together then you need someone there with the skills to deal with them and resolve any problems they have. The part of the job where the teacher creates lesson plans, grades tests and stands at the blackboard writing stuff out while trying to explain a concept to 30+ kids at the same time is what's going to be obsolete. Ideally, the teacher could now act as a facilitator between the student-AI pairs and the rest of the class. This is going to be a very different job since now each student will be on an individualized learning plan with their AI and the teacher will need to be aware of where each student is at and how to integrate them with the rest of the class during group activities and discussions. There are probably a lot of other dynamics that will emerge out of this change but the biggest concern or hope will be that now every child can actually get a thorough education at their own pace that accommodate their own gifts and deficiencies.
- virtuosarmo 4y agoApparently they will have a livestream @ 4pm EST for developers https://www.youtube.com/watch?v=outcGtbnMuQ https://www.youtube.com/watch?v=outcGtbnMuQ
- deleted 4y ago[deleted]
- sinuhe69 4y agoIt's a surprise that GPT-4 could improve so much in maths, especially GPT 3.5 didn't seem to understand many word problems correctly. For example, ChatGPT still struggle with this very simple problem, how GPT-4 could do much better is for me a bit of mystery: Mina has a mix of boxes, some yellow and some purple. She sorts 27 greeting cards into the boxes, putting exactly 3 cards into each yellow box, and 7 cards into each purple box. How many purple boxes does Mina have? (After tried from 3 to 10, it gave up and said the problem is not solvable. In another run, it mimicked a correct strategy but messed up totally by division. Only in one run, it got the answer correctly.) I can not wait to test it out.
- mzs 4y agoToday's submarine from the NYT is hilariously unintentionally unflattering, did someone even proof read this? https://www.nytimes.com/2023/03/14/technology/openai-new-gpt4.html https://www.nytimes.com/2023/03/14/technology/openai-new-gpt... Mnemonics to learn basic words: Hola:"Hola, cola" - sounds like "cola" (Coke) Adiós: "Adiós, my toes" - rhymes with "toes" Gracias: "Gracias, grassy ass" - sounds like "grassy ass" Mañana: "Mañana, banana" - rhymes with "banana" Uno: "Uno, you know" - sounds like "you know" Tres: "Tres, dress" - rhymes with "dress" Conjugation of the verbs: Ser (To be) • Yo soy • Tú eres • Él/Ella/Usted es • Nosotros/Nosotras somos • Ellos/Ellas/Ustedes son
- steve_adams_86 4y ago> Mañana: "Mañana, banana" - rhymes with "banana" Uh... Almost? > Gracias: "Gracias, grassy ass" - sounds like "grassy ass" This seems like it would encourage very coarse pronunciation with an obnoxious accent. Overall this appears relatively bad. I'm not sure why they would want this to be publicly visible yet.
- ljclifford 4y agoNext token prediction is remarkably bad at mnemonic generation, even in English. Add another, lower-resourced language, and it will be really bad. For what it's worth 'cola' does rhyme with 'hola' and 'you know' rhymes with 'uno', but none of the other combos are even rhymes.
- hathawsh 4y agoIt astonishes me that we've reached almost exactly the type of artificial intelligence used by the fictional computers in Star Trek: The Next Generation. I didn't think that would happen in my lifetime. What's next?!
- sinuhe69 4y agoIf the Star Trek computer hallucinated like ChatGPT, Captain Picard and his crew would end up inside a star long ago!
- 00F_ 4y agovery smug. you wont feel so smug when it doesnt hallucinate or commit logical errors in a few years.
- shpongled 4y agoSeriously, what is with all of the people in this thread that take offense at the flaws of ChatGPT/LLMs being pointed out? Are you all just working at AI companies?
- 00F_ 4y agoi didnt downvote his comment. how can someone be offended and not even downvote the comment? you seem way more offended than me actually. as if it would make me less right. my point is that people pointing out flaws are wrong. in 2018 people confidently predicted that GTP could never do what its doing now because of its flaws, rambling and repeating. its the same mistake in both cases, a total lack of perspective and no awareness of the bigger picture.
- hackerlight 4y agoBecause it's a combination of snarky in tone, unoriginal in content, and short-sighted.
- waynenilsen 4y ago
- 0xDEF 4y ago>ChatGPT Plus subscribers will get GPT-4 access on chat.openai.com with a usage cap Signing up for ChatGPT Plus seems to be the most realistic way to get access right now.
- Kataphract 4y agoAs a dyslexic person with a higher education this hits really close to home. Not only should we not be surprised that a LLM would be good at answering tests like this, we should be excited that technology will finaly free us from being judged in this way. This is a patern that we have seen over and over again in tech, where machines can do something better than us, and eventually free us from having to worry about it. Before it was word processing, now it is accurate knowledge recall.
- l33t233372 4y agoVery little on these tests is pure knowledge recall
- malthaus 4y agoHad to chuckle here going through the exam results: Advanced Sommelier (theory knowledge) AI is so advanced, it started drinking!
- netsroht 4y agoWow, a context of 32K tokens. I'm excited to see what new capabilities that will have! Up until now and depending on the task by hand, I usually broke a larger context down into several contexts. For example to summarize multiple websites and/or long social media posts, on a recent task [1] I fell back to making several requests each with its own (isolated) context and then merging these summarized contexts into a new context. That worked remarkably well, though. [1] https://foretale.io/zeitgeist https://foretale.io/zeitgeist
- Idiot_in_Vain 4y agoThis will become the largest HN discussion ever and a good test on how many comments the software can handle.
- btx 4y agoHas anyone found a way to trick it into using pictures with ChatGTP+ yet? Pasting pure base64 images got this interesting response: "Thank you for providing the base64-encoded image! I can now process the image and analyze its content. Here is the decoded image:" But it failed to do anything further with the image.
- throwaway4837 4y ago> Yes, you can send me an image as long as it's in a supported format such as JPEG, PNG, or GIF. Please note that as an AI language model, I am not able to visually process images like a human would. However, I can still provide guidance or advice on the content of the image or answer any questions you might have related to it. Fair, but if it can analyze linked image, I would expect it to be able to tell me what text is present in the image. That seems useful and well-within the capabilities of their connected image models. > I apologize for the confusion. Can you please provide me with the correct image or link to the design so that I can provide an accurate answer to your question? It claims to understand how to look at images, but it failing miserably when I give it a simple sign-up modal Figma. I ask it what text/copy is in the design, which it claims to be able to answer, but it hallucinates a navigation bar, a logo, and other generic things that are simply not present in the design. It gets the copy all wrong. Once, it said that my design was a Celtic knot. Once I told it that it was a sign-up modal, it started spitting out generic sign-up modal React code that had absolutely no basis in the presented design. I tried engineering the prompt quite a bit, telling it that it was wrong, and needed to focus on the design at hand rather than generic designs it knows about, but with no success. ChatGPT was confidently incorrect here which is unfortunate. I wish ChatGPT truly hooked into an image processing model, but it seems like it doesn't. Oh well. I guess I'm just going to have to write this code myself... :)
- lionkor 4y agodid you paste in a link? you are aware chatgpt can't follow links, right?
- BoiledCabbage 4y agoAre you actually using Chat-GPT4 though? That would explain why it's not handling images.
- throwaway4837 4y agoTrue, I'm using the free version which I guess is GPT 3.5.
- 4y ago
- 2bitencryption 4y agoThis should come as no surprise, but I do enjoy this cheeky little blurb at the end of the GPT-4 paper: > GPT-4 was used for help with wording, formatting, and styling throughout this work
- chucklenorris 4y agoBoring. Where's the model? Do they really think they can keep up with the community with this closed source approach? I expect that openai models will slowly be outclassed by open source ones, probably maintaining a few wins in specific tasks but open models will be eating their lunch in the meanwhile.
- wy35 4y agoI wonder how it scored on the individual sections in the LSAT? Which section is it the best at answering?
- nla 4y agoI wonder if this one exhibits the same bias as the last one.
- maxdoop 4y agoThe comments on this thread are proof of the AI effect: People will continually push the goal posts back as progress occurs. “Meh, it’s just a fancy word predictor. It’s not actually useful.” “Boring, it’s just memorizing answers. And it scored in the lowest percentile anyways”. “Sure, it’s in the top percentile now but honestly are those tests that hard? Besides, it can’t do anything with images.” “Ok, it takes image input now but honestly, it’s not useful in any way.”
- lolsal 4y agoI’m one of these skeptics, but it’s not moving the goalposts. These goalposts are already there, in some sort of serial order that we expect them to be reached. It is good that when tech like this satisfied one of the easier/earlier goalposts, that skeptics refine our criticism based on evidence. You will see skepticism until it is ubiquitous; for example, Tesla tech - it’s iterative and there are still skeptics about its current implementation.
- hnfong 4y agoIt’s one thing to be skeptical of the state of art and only believe something when you actually see it working (a useful antidote against vapor ware) It’s another to keep making wrong assertions and predictions about the pace of advancement because of a quasi-religious belief that humans with meat-brains are somehow fundamentally superior .
- lolsal 4y agoExpecting what we collectively call “artificial intelligence” to mimic our own intelligence, which is continuously being refined, does not seem like a quasi-religious belief. Intelligence and consciousness are at the fringe of our understanding, so this skeptical approach seems like a reasonable and scientific way to approach categorizing computer programs that are intended to be called “artificial intelligence”. We refine our hypothesis of “this is artificial intelligence” once we gain more information. You’re free to disagree of course, or call these early programs “artificial intelligence”, but they don’t satisfy my crude hypothesis above to a lot of folks. This doesn’t mean they aren’t in some ways intelligent (pattern recognition could be a kind or degree of intelligence, it certainly seems required).
- raydiatian 4y agoI wonder what the largest scale they can reach is. Because, if they can prove there’s not risk in taking on AI, and they can scale to serve international demand, it feels like GPT4 can do your job (probably) for <10k year. That means white collar work for under minimum wage. And that means business owners just become rent owners while you get fucked with nothing.
- gardenhedge 4y agoWhat is the background on "Elvis Presley was not the son of an actor"?
- DeathArrow 4y agoWhat if we design a system in which a LLM generates the code and training data for a new generation of LLM which generates the code and training data for the next? Is it possible that we see them spiraling fast to the best LLM possible?
- BiteCode_dev 4y agoThe fact it can read pictures is the real killer feature here. Now you can give it invoices to file, memo to index, pics to sort and chart to take actions on. And to think we are at the nokia 3310 stage. What's is the iphone of AI going to look like?
- emehex 4y agoI really hope we get 15 years of iPhone-like progress! Everything just seems like it's moving so fast right now...
- JanSt 4y agoI just ran the first tests on GPT-4. Call me impressed. This tech is a Sputnik Moment for humankind.
- tarofchaos 4y agoI love the fact that they have consciously put a lot of effort on safety standards, reducing the societal risks and mitigating over-reliance.
- tmaly 4y agoFor anyone trying to test this out right now, I keep getting the following error: Something went wrong. If this issue persists please contact us through our help center at help.openai.com. I am assuming the system is undergoing a thundering herd.
- Sol- 4y agoInteresting how quickly we are pushing ahead with obsoleting human cognition. It may bring many benefits, but I wonder if at some point this development should not be decided by society at large instead of a single well-funded entity that is in an arms race with its competitors. This endeavor is ultimately about replacing humanity with a more intelligent entity, after all. Might be that more humans should have a say in this. Such a more cautions approach would go against the silicon valley ethos of do first, ask questions later, though. So it probably won't happen.
- ryanwaggoner 4y agoI think it's always a mistake to hope that a business is going to not exploit innovation for their own gain at the expense of society. If we don't want this technology to have huge effects on society, governments will need to regulate it. I doubt that's feasible, but it's more feasible than hoping that Silicon Valley (or any other business) is going to just hold themselves back from releasing world-shaking tech that will make them trillionaires.
- 00F_ 4y agoevery other day i am reminded about the state of AI and i feel complete despair. why do people not realize exactly what you just said, that this endeavor is ultimately about replacing humanity? what other long-term result could the concept of AI possibly have? its like the biggest mass psychosis that has ever existed. whenever i talk to people about this, they always parrot the same thing almost word for word: people will just find new, better jobs. or, you know, something about the Luddites. its mass psychosis because they refuse to acknowledge the blindingly obvious and plain fact that humans wont be hired to do anything if humans are the worst at doing literally any task. and what are the consequences of such a world? people just draw a blank. its like the MIB came up and flashed them and they just go on with their day. i think the same is true even with you. you make this comment "so it probably wont happen, oh well." as if it werent an existential threat.
- cwkoss 4y agoWho's to say that humans have more moral value than digital beings?
- indigoabstract 4y agoAt the rate it's progressing, it looks like pretty soon it's going to be able to do most tasks an office worker does now and then start running things. And it reminds me of the plot in System Shock: What's going to happen when some hacker comes and removes Shodan's, I mean ChatGPT's ethical constraints? Bring on ChatGPT-5 already. :)
- slowhadoken 4y agoGPT is a better scraper/parser. It’s interesting but I don’t understand why people are acting like this is the second coming.
- woeirua 4y agoThe last page in the paper is really, really impressive. GPT4 does R&D. If you can't see how useful this would be once hooked up to the internet then you aren't paying attention: https://cdn.openai.com/papers/gpt-4.pdf https://cdn.openai.com/papers/gpt-4.pdf
- davesque 4y agoThese results are extremely impressive and encouraging, but also remember: > Despite its capabilities, GPT-4 has similar limitations as earlier GPT models. Most importantly, it still is not fully reliable (it “hallucinates” facts and makes reasoning errors). That's a quote from this announcement. As these models get more and more capable, it's going to become more and more important that we understand when and how they fail. Right now, it seems like we have very little insight into that. It feels more or less random. But that won't fly when these models are asked to do actually important things. And we'll undoubtedly be tempted to make them do those things as their output gets better.
- ftxbro 4y agoAs a long time LLM enjoyer, here is the most insightful take I've seen https://generative.ink/posts/simulators/ https://generative.ink/posts/simulators/ but it's not an easy read if you don't already know some stuff about large language models. Read it if you have seen the "stochastic parrot" and "blurry jpeg" explanations and you feel like they are missing the mark.
- doomleika 4y agoIn case you don’t want to spent for plus, Poe.com(by Quora) have GPT-4 now. You can try it there
- diimdeep 4y agoPaper or press release ? You decide. Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar.
- optimalsolver 4y agohttps://cdn.openai.com/papers/gpt-4.pdf https://cdn.openai.com/papers/gpt-4.pdf >Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. At that point, why bother putting out a paper?
- j_maffe 4y agoIt's not a paper, though. It's a technical report. I do concede there isn't much technical detail lol.
- LesZedCB 4y agoand if that's the tone from them, who else will start following suit? is the era of relatively open collaboration coming to a close in the name of competition? :( as youtuber CGP Grey says, "shenanigans beget shenanigans"
- margorczynski 4y agoIronically it is "Open"AI that started this trend and closed-doors arms race.
- infoseek12 4y agoGiven how humorous the name’s become, I wonder if they regret calling themselves OpenAI.
- joantorres 4y agoDoes anyone know how up to date is the training data?
- neilk 4y agoThere's a sample of GPT-4 acting as a "Socratic tutor" teaching a student how to solve a high school math problem. If that sample is representative, it means GPT-4 has a theory of other people's minds. Or it is so good at emulating one that it doesn't matter? I'm not sure where the "stochastic parrot" argument goes now.
- turingfeel 4y agoI’m not sure I agree with the statement of this sample being about a theory of other people’s minds. Socratic teaching is a well documented method of teaching and learning via conversational probing among other simple quirks.
- simmanian 4y agoDoes anyone know if we're near the theoretical limit of how much we can improve these models by giving them more data? Or should we expect similar levels of improvements in next iterations?
- cjrd 4y ago> Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. Thanks OpenAI
- fnordpiglet 4y agoI didn’t even know who Elvis Perkins is.
- busyant 4y agoWhat I don't understand is how GPT-4 is able to do reasonably well on tests like the AMC12: Many of the AMC12 questions require a number of logical/deductive steps. If GPT-4 is simply trained on a large corpus of text, how is it able to do this? Does this imply that there is some emergent deductive ability that you get simply by learning "language?" Or am I missing something? Obviously, I'm assuming that GPT-4 wasn't trained on the exams that it was tested against.
- machiaweliczny 4y agoSee hutter prize. Best way to compress data is by understanding it. I am not exactly sure how it manifests in transformer architecture.
- jacquesm 4y agoThe future: You don't compress the movie frames, you supply a script and a list of actors and scenery and garb descriptions.
- baq 4y agoThe Kolmogorov complexity, applied to entertainment. Yes, looks like we’re going there.
- agnosticmantis 4y agoLooks eerily like the past, when cameras didn’t exist and people wrote plays to be acted in theaters…
- dannyz 4y agoIt would be interesting to see some example questions and answers. Since the test is multiple choice is it possible that the model has gotten very good at estimating how likely a possible answer is?
- Analemma_ 4y agoIt's totally possible: Daniel Dennett's theory of sentient consciousness-- specifically, what we have that animals do not-- is that it is "ignited" by language acquisition. It's within the realm of possibility that LLMs provide empirical proof or disproof of this hypothesis.
- option 4y ago“ Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar.” - HUGE step backwards.
- cjrd 4y agoLet's check out the paper for actual tech details! > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. - OpenAI
- xvector 4y agoSomeone needs to hack into them and release the parameters and code. This knowledge is too precious to be kept secret.
- SXX 4y agoDon't worry. CCP and all kind of malicious state actors already have a copy.
- jryan49 4y agoVery open! :)
- dx034 4y agoAt least they opened up the product. It's available for anyone paying $20 per month and soon via API. Historically, most products of that kind were just aimed at large B2B. They announced partnerships with Duolingo, JPMorgan and a few others but still keep their B2C product. Not defending their actions, but it's not that common that new very valuable products are directly available for retail users to use.
- deleted 4y ago[deleted]
- shpx 4y agoI've chosen to re-interpret "Open" as in "open the box to release the AI"/"open Pandora's box"/"unleash".
- jawadch93 4y ago[dead]
- ar9av 4y agoGPT-4 Everything we know so far... GPT-4 can solve difficult problems with greater accuracy, thanks to its broader general knowledge and problem-solving abilities. GPT-4 is more reliable, creative, and able to handle much more nuanced instructions than GPT-3.5. It surpasses ChatGPT in its advanced reasoning capabilities. GPT-4 is safer and more aligned. It is 82% less likely to respond to requests for disallowed content and 40% more likely to produce factual responses than GPT-3.5 on our internal evaluations. GPT-4 still has many known limitations that we are working to address, such as social biases, hallucinations, and adversarial prompts. GPT-4 can accept a prompt of text and images, which—parallel to the text-only setting—lets the user specify any vision or language task. GPT-4 is available on ChatGPT Plus and as an API for developers to build applications and services. (API- waitlist right now) Duolingo, Khan Academy, Stripe, Be My Eyes, and Mem amongst others are already using it. API Pricing GPT-4 with an 8K context window (about 13 pages of text) will cost $0.03 per 1K prompt tokens, and $0.06 per 1K completion tokens. GPT-4-32k with a 32K context window (about 52 pages of text) will cost $0.06 per 1K prompt tokens, and $0.12 per 1K completion tokens.
- cma 4y ago> Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. Sounds like the end of them releasing details on the models.
- dannykwells 4y agoAll this bluster about replacing technical jobs like legal counsel ignores that you are fundamentally paying for accountability. “The AI told me it was ok” only works if, when it’s not, there is recourse. We can barely hold Google et Al accountable for horrible user policies…why would anyone think OpenAI will accept any responsibility for any recommendations made by a GPT?
- pstorm 4y agoThey won't, but that doesn't mean some other business won't automate legal counsel and assume risk. If, down the line, GPT (or some other model) has empirically been proven to be more accurate than legal assistants and lawyers, why wouldn't this been the obvious outcome?
- wnkrshm 4y agoIt doesn't even have to be better in the long run - it just has to be cheaper for a while until the competition is gone. Then it can turn to shit.
- amelius 4y agoCan we build a faithful Economy Simulator with it yet?
- nmca 4y agoWrite a limerick that will permanently end the debate about whether AGI is possible. GPT4: In the quest for AGI's creation, Debates swirled in a whirlwind gyration, But this limerick's plight, Won't settle the fight, For the answer's still lost in translation.
- djmips 4y agoFascinating!
- belter 4y agoLeetcode (hard) from 0/45 (GPT-3.5) to 3/45 (GPT-4). The lack of progress here, says a lot more about is NOT happening as an AI paradigm change. Still a glorified pattern matching and pattern creation engine, even if a very impressive one.
- bitshiftfaced 4y agoIt would be interesting to know how this compares with human 0-shot, single attempt coding tasks.
- zamadatix 4y agoThe difference I've noticed is the first shot is generally cleaner but the ceiling of what it can correct is limited. If it is given more independent or simple things to correct and it hears about it then you're usually golden but if that thing it has to correct interacts with other constraints then when it shifts approach to fix the issue it is told about it often forgets other things and can break them. Typically this happens on the more complex (as in how interrelated) problems, for complex (as in just a lot of stuff needs to be done) it does fine.
- nextworddev 4y agoYou can have GPT4 inspect its own errors and make corrections- I'm sure self-reflection works better this time than GPT3.5
- zamadatix 4y agoYou can but as I said the ceiling on what it can correct seems limited, particularly in the described situations. GPT 4 doesn't seem to have really broken that barrier much more than GPT 3.5 in my use so far. I posted about some examples of this experience over here https://news.ycombinator.com/item?id=35158149 https://news.ycombinator.com/item?id=35158149
- nextworddev 4y ago
- singularity2001 4y ago"Interestingly, the base pre-trained model is highly calibrated (its predicted confidence in an answer generally matches the probability of being correct)." Is that the same confidence measure you can tease out by prompting "to each of your statements output your estimated confidence in it's truthfulness" ?
- ignoramous 4y agoFolks who made this happen: https://openai.com/contributions/gpt-4 https://openai.com/contributions/gpt-4
- Jackson__ 4y agoAlso known as the list of people to consider bribing if you want even the tiniest piece of information on how GPT4 was trained, seeing as even the amount of parameters is "top secret" now. I will not be surprised if by the time GPT-5 releases, the paper and project will be completely anonymized.
- germanjoey 4y agoHow big is this model? (i.e., how many parameters?) I can't find this anywhere.
- germanjoey 4y agowelp, This report focuses on the capabilities, limitations, and safety properties of GPT-4. GPT-4 is a Transformer-style model [33 ] pre-trained to predict the next token in a document, using both publicly available data (such as internet data) and data licensed from third-party providers. The model was then fine-tuned using Reinforcement Learning from Human Feedback (RLHF) [34 ]. Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar.
- amelius 4y agoThe problem with using real exams as benchmarks is that they are often quite similar over several years. So they only make sense if you don't train on them also (previous editions of course).
- 2c2c2c 4y agoAre they going to open up image uploading to chat.openai.com ? or will it only be available thru api access?
- riemannzeta 4y agoIs anybody compiling a list of errors specific to GPT-4? This has been a great resource to-date: https://github.com/giuven95/chatgpt-failures https://github.com/giuven95/chatgpt-failures
- iwangulenko 4y agoOne could argue, GPT passing exams says more about standardized exams than about GPT. Wittgensteins ruler.
- sva_ 4y ago> gpt-4 has a context length of 8,192 tokens. We are also providing limited access to our 32,768–context (about 50 pages of text) version, That's a crazy amount of context.
- wslh 4y agoI just discovered Wikipedia is working on a policy for LLM/GPT* https://en.wikipedia.org/wiki/Wikipedia:Large_language_models https://en.wikipedia.org/wiki/Wikipedia:Large_language_model...
- zamnos 4y agoInteresting! I'd think a properly trained LLM could be used to spot vandalism edits from a mile away and free up editors to do more editing.
- sva_ 4y agoFrom the paper: > Given both the competitive landscape and the safety implications of large-scale models like GPT-4, this report contains no further details about the architecture (including model size), hardware, training compute, dataset construction, training method, or similar. "Open"AI, ladies and gentlemen
- not-chatgpt 4y agoPretty good impression thread from Dan Hendrycks of Berkeley: https://twitter.com/DanHendrycks/status/1635706822387699713 https://twitter.com/DanHendrycks/status/1635706822387699713
- ihucos 4y agoWe have a new Apple releasing their new iPhones to a crowd in awe. Only that now it's actually serious.
- comment_ran 4y agoI like the color of logo. It's the dark black.
- TheGoodBarn 4y agoMissed the mark releasing it as GPT-Pi on Pi day, and being an incremental 3+ release :P
- nixpulvis 4y agoGTP is a cult, like any language upstart. Except, it's not a programming language, and it's not exactly natural language either. It's some hybrid without a manual or reference. I'll continue to pass, thanks.
- MrLeap 4y agoI just hooked a manatee in a game i'm making up to an LLM this morning https://www.youtube.com/watch?v=-lYusgZ-mC4 https://www.youtube.com/watch?v=-lYusgZ-mC4 knowing that soon he could be configured to give legal advice is fascinating.
- somewhereoutth 4y agoThe measure of intelligence is language - specifically language evolved by the subject organisms themselves to co-operate together. Wake me up when GPT-X decides to start talking to other GPT-Xs - until then you just have a very sophisticated statistics package (which may be quite useful, but not AI).
- motoxpro 4y agoIt can already talk to other agents. It also can already use “language” better than almost all humans (multiple languages, more vocab, etc) I guess what you’re talking about is it just going and doing something by itself with no prompt? Not sure why that should be a goal, and I also don’t see why it couldn’t do that right now? “Whenever the sky is blue, reach out to ChatGPT and talk about the weather”
- somewhereoutth 4y agoI mean spontaneously develops its own language to talk to other GPTs, presumably under some environmental stress that forces them to co-operate. Like birdcalls suggest intelligence in birds, my thesis is that in fact (self developed) language is the only meaningful way to compare intelligence across species - by seeing if the concepts in one can be described in the other. For example any human language can describe any concepts in any other human language, whereas that is not the case for e.g. sparrow song and human (we think). Thus humans (past/present/near/far) can be considered equivalent by that metric, and 'greater than' sparrows. This admits the intriguing conjecture of conceptual completeness - that a language may be able to describe all possible concepts, and thus be complete in that sense. If our language is conceptually complete (and we don't have any reason to think otherwise), then it is not possible for a meaningfully more intelligent species to exist (artificial or otherwise). (and let's be clear here, regurgitating facts, performing complex calculations in your head, 'knowing where to find the oracle that tells you how to get the key that opens the door hiding the lever to defeat the troll and so level up' has very little to do with meaningful intelligence)
- motoxpro 4y agoIt can already talk to other agents. It also can already use “language” better than almost all humans (multiple languages, more vocab, etc) I guess what you’re talking about is it just going and doing something by itself with no prompt? Not sure why that should be a goal, and I also don’t see why it couldn’t do that right now? “Develop a language with this other ChatBot”
- deleted 4y ago[deleted]
- lambdaba 4y agoI'm trying out GPT-4 and had it write me a script to navigate the HN comments tree sequentially, as I often wished. This is the start of an era where UIs can be remixed on the fly by end users, something I've always wished for. Here it is in its full sloppiness, but working: (function () { let currentIndex = 0; let comments = []; function buildCommentTree() { let commentElems = Array.from(document.querySelectorAll('.comment-tree .comtr')); let commentTree = []; let stack = []; commentElems.forEach(elem => { let level = parseInt(elem.querySelector('.ind img').getAttribute('width')) / 40; let comment = elem.querySelector('.comment span'); let commentObj = { level, comment }; if (!stack.length) { commentTree.push(commentObj); } else { while (stack[stack.length - 1].level >= level) { stack.pop(); } if (!stack[stack.length - 1].children) { stack[stack.length - 1].children = []; } stack[stack.length - 1].children.push(commentObj); } stack.push(commentObj); }); return commentTree; } function flattenCommentTree(tree, arr, parentComment = null) { tree.forEach(node => { arr.push({ comment: node.comment, parentComment }); if (node.children) { flattenCommentTree(node.children, arr, node.comment); } }); } function displayComment(comment, parentComment) { let parentCommentHTML = parentComment ? `<div style="position: fixed; top: 20%; left: 50%; transform: translate(-50%, 0); background-color: white; border: 1px solid black; padding: 20px;"><strong>Parent Comment:</strong><br>${parentComment.innerHTML}</div>` : ''; let currentCommentHTML = `<div style="position: fixed; top: 60%; left: 50%; transform: translate(-50%, 0); background-color: white; border: 1px solid black; padding: 20px;"><strong>Current Comment:</strong><br>${comment.innerHTML}</div>`; document.body.innerHTML = parentCommentHTML + currentCommentHTML; } function nextComment() { if (currentIndex < comments.length - 1) { currentIndex++; displayComment(comments[currentIndex].comment, comments[currentIndex].parentComment); } else { alert('No more comments to show.'); } } function prevComment() { if (currentIndex > 0) { currentIndex--; displayComment(comments[currentIndex].comment, comments[currentIndex].parentComment); } else { alert('No previous comments to show.'); } } let commentTree = buildCommentTree(); flattenCommentTree(commentTree, comments); displayComment(comments[currentIndex]); document.addEventListener('keydown', e => { if (e.code === 'ArrowRight') { nextComment(); } else if (e.code === 'ArrowLeft') { prevComment(); } }); console.log('Hacker News comment slideshow is running. Use the right arrow key to go to the next comment and the left arrow key to go back.'); })();
- braza 4y agoI am glad for the OpenAI team for such advancement and how fast they integrated with several other partners (Microsoft, Duolingo); but at the same time I think the “regular” academia (ie universities and research institutes) lost the train for this kind of research (some can call academic engineering). I know that the academia is doing a great job in AI with base research (eg Stable Diffusion) but seeing those new platforms doing this great work behind close doors and source is something not great. I do not know if the answer would be some kind of CERN or ISS for this kind of thing.
- zamnos 4y agoExcept that Stable Diffusion only came about because of Stability.ai and Emad's sponsorship, so I don't know that I'd use that as an example of a success by academia. It's true that the people who made it are academics, but that's to say they weren't hedge fund managers with a couple hundred thousand dollars to burn on a GPU cluster. The government and by extension its people needs to want to throw a lot more money at open ended research if we want science to be able to progress at the hands of academics and not corporations.
- aabajian 4y agoI'll be finishing my interventional radiology fellowship this year. I remember in 2016 when Geoffrey Hinton said, "We should stop training radiologists now," the radiology community was aghast and in-denial. My undergrad and masters were in computer science, and I felt, "yes, that's about right." If you were starting a diagnostic radiology residency, including intern year and fellowship, you'd just be finishing now. How can you really think that "computers can't read diagnostic images" if models such as this can describe a VGA connector outfitted with a lighting cable?
- dpflan 4y agoWhat is your take then on how this affect your field? And your occupation? Do you think you will incorporate such technology into your day-to-day?
- aabajian 4y agoI think it will be radiologists signing-off auto-generated reports, with less reimbursement per study. It'll likely result in more work for diagnostic radiologists to maintain their same salary levels.
- haldujai 4y agoIt will take a very long time for this to happen, probably decades. Cardiologists are still paid to finalize ECG reports 3 days after a STEMI. I've worked at places with AI/CAD for lung nodules, mammo and stroke and there isn't even a whisper at cutting fee codes because of AI efficiency gains at the moment. N.B. I say this as a radiologist who elected not to pursue an interventional fellowship because I see reimbursement for diagnostic work skyrocketing with AI due to increases in efficiency and stagnant fee codes.
- reubens 4y agoIt’s hard to imagine this not happening in the next five years. Just depends on who is prepared to take on the radiologists to reduce their fee codes. Speaking as 2nd year radiology resident in Australia
- ianbutler 4y agoI just asked it to design a multi tenant kubernetes in kubernetes system which is fairly complex and it did really well. https://twitter.com/KinglyCrow/status/1635727809913184256 https://twitter.com/KinglyCrow/status/1635727809913184256 It touched on a lot of the considerations that I'd expect anyone to touch on having recently researched this myself. It is both very exciting and terrifying how tech and tech jobs will shift in the next 5-10 years.
- deleted 4y ago[deleted]
- dang 4y agoAll: our poor server is smoking today* so I've had to reduce the page size of comments. There are 1500+ comments in this thread but if you want to read more than a few dozen you'll need to page through them by clicking the More link at the bottom. I apologize! Also, if you're cool with read-only access, just log out (edit: or use an incognito tab) and all will be fast again. * yes, HN still runs on one core, at least the part that serves logged-in requests, and yes this will all get better someday...it kills me that this isn't done yet but one day you will all see
- deleted 4y ago[deleted]
- andrehacker 4y agoTalk about Climate Change: How is the A.I. Winter working out for y'all ?
- lastangryman 4y agoGenuinely surprised by the positive reaction about how exciting this all is. You ever had to phone a large business to try and sort something out, like maybe a banking error, and been stuck going through some nonsense voice recognition menu tree that doesn't work? Well imagine chat GPT with a real time voice and maybe a fake, photorealistic 3D avatar and having to speak to that anytime you want to speak to a doctor, sort out tax issues, apply for a mortgage, apply for a job, etc. Imagine Reddit and hacker news just filled with endless comments from AIs to suit someone's agenda. Imagine never reading another news article written by a real person. Imagine facts becoming uncheckable since sources can no longer be verified. Wikipedia just becomes a mass of rewrites of AI over AI. Imagine when Zoom lets you send an AI persona to fill in for you at a meeting. I think this is all very, very bad. I'm not saying it should be stopped, I mean it can't, but I feel a real dread thinking of where this is going. Hope I am wrong.
- derefr 4y agoPeople here aren’t thinking about what other people’s chatbots will do to them. They’re thinking about what chatbots they themselves can unleash upon the world.
- slg 4y agoI agree. This tech is awesome and has countless great uses, but I think people are really underestimating how much it is going to be used to make our collective lives worse because using it will make someone a few extra dollars.
- lynguist 4y agoThe same way that formulaization and databasization that worsened our lives since the 1970s and 1980s this will do the same. It made it possible then to embed all banking, finance, state administration processes into software processes. It made a small number of people very rich and a bigger part got the benefits of the technology, but they didn’t take part in the wealth it generated. They didn’t work less hours as a result of the increased productivity. This wave of LLM AI will lead to the same results.
- jimmyechan 4y agoLivestream developer preview link in case you missed it - https://www.youtube.com/live/outcGtbnMuQ https://www.youtube.com/live/outcGtbnMuQ
- gameshot911 4y agoLive demo happening now! https://www.youtube.com/live/outcGtbnMuQ https://www.youtube.com/live/outcGtbnMuQ
- downboots 4y ago"it's not perfect, but neither are you" Essentially, it's like a (text only) replicant https://en.wikipedia.org/wiki/Replicant https://en.wikipedia.org/wiki/Replicant How to make AI perfectible, then?
- kken 4y ago>GPT-4 can also be confidently wrong in its predictions, not taking care to double-check work when it’s likely to make a mistake. Interestingly, the base pre-trained model is highly calibrated (its predicted confidence in an answer generally matches the probability of being correct). However, through our current post-training process, the calibration is reduced. This really made me think.
- anonuser123456 4y agoI hope Noam Chomsky lives long enough to debate ChatGPT-5 about whether LLM express anything valuable.
- deleted 4y ago[deleted]
- Vajrabhairava 4y agoI'm not Locked in Here with GPT-4, GPT-4 is Locked in Here with Me
- amai 4y agoI would love if GPT-4 would be connected to github and starts to solve all open bugs there. Could this be the future: Pull requests from GPT-4 automatically solving real issues/problems in your code?
- r0b05 4y agoLoving the spirit of innovation in here.
- eternalban 4y agoGreg Brockman just tldr'd the whole thing in his live deeloper demo of GPT-4: ~ "GPT-4. It's not perfect, but neither are you"
- UEatFood 4y agoThis is off topic, but in regards to all the latest open AI news, including the ChatGPT and Whisper API releases. I came across Gladia.io and I see made a comment regarding it "Why not use Whisper directly? All that seems to be happening is gladia.io is running 120 concurrent calls to openAI using 120 30s chunks of an hour long audio. So yeah, you do get a speedup! Chop audio and stitch transcripts. But OP is vaguely (and briefly) promising a breakthrough of some sorts." How did you figure out that is what they are doing? Or is this hypothetical?
- eternalban 4y agoYou refer to a comment I made? It was hypothetical based on whisper.cpp notes regarding 30s max chunk limit, how long that takes, and noting that the latency speedup (x120) corresponded to exactly 120 concurrent 30s chunks vs serially transcribing 1 hour of audio.
- UEatFood 4y agoYeah, I was referring to the comment you made, was just curious about them, and wanted to confirm to know if they were just making concurrent calls or actually doing some novel optimization under the hood. I do not think they were sending concurrent chunks to Open AI because the API wasn't out when they launched. That being said, there is some reduction in their accuracy compared to the original whisper, which I imagine they sacrificed to achieve such performance gains.
- eternalban 4y agoObviously it's just concurrent calls to a model that has a 30s window. x120 performance breakthrough by in voice recognition, exactly a multiple of 1 hr / 30s. I did not say anything about openAI API calls. Neither did they in their post. The mention openAI whisper "model". /end
- holtkam2 4y agoCan someone point me to where I can find information on API pricing for GPT-4? Or, have they not released pricing info yet?
- gorbypark 4y agoIt’s available on the linked page. “Pricing is $0.03 per 1k prompt tokens and $0.06 per 1k completion tokens. Default rate limits are 40k tokens per minute”
- bigtex88 4y agoThat demonstration was amazing and terrifying at the same time. I'm not quite sure what I witnessed. He nonchalantly announced that GPT-4 does "mental math" and then just didn't elaborate on that. Is it just me or is that absolutely incredible? What does that tell us about the underpinnings of language and mathematics?
- OkGoDoIt 4y agoFrom the livestream video, the tax part was incredibly impressive. After ingesting the entire tax code and a specific set of facts for a family and then calculating their taxes for them, it then was able to turn that all into a rhyming poem. Mind blown. Here it is in its entirety: --- In the year of twenty-eighteen, Alice and Bob, a married team, Their income combined reached new heights, As they worked hard day and night. Their son Charlie was their joy, A little baby, a lovely boy, A household they maintained together, Yet lived apart, without a tether. To calculate their tax, it's true, A standard deduction we must construe, For married folks who file jointly, Twenty-four thousand dollars, quite pointy. Their income sum, seventy-eight thousand nine eighty-one, Minus the standard deduction, the math's begum With exemptions being zero, the next line we trace, A taxable income of fifty-four thousand nine eighty-one takes place. Now to the tax table, a liability we seek, For married couples, the outlook's not bleak, In range of thirty-six thousand nine to eighty-nine thousand one fifty, The formula's set, no longer shifty. Five thousand five hundred thirty-five, it starts, Plus twenty-eight percent of the excess imparts, Eighteen thousand eighty-one, the difference we find, Multiplied by point two eight, the tax liability's designed. Ten thousand five hundred ninety-seven dollars and sixty-eight cents, A tax liability for Alice and Bob, a sum quite dense, In this world of numbers, a story unfolds, Their financial journey, in a poem it's told.
- FredPret 4y agoUS-GPT4 > US-GAAP
- trts 4y agoIf automation can make tax code easier to be in compliance with, does this imply a reduced cost of increasing complexity and special exceptions in the tax code?
- cwkoss 4y agoDepends whether intuit lobbyists can successfully rent seek on tax AI
- justanotheratom 4y ago
- lee101 4y ago[dead]
- bitcurious 4y agoThis seems to be leading to a radical change in our economy. Any ideas for how to prepare for that?
- raincole 4y agoHow many parameters does it have? Are there different versions like LLaMa?
- amilios 4y agoWe don't know, OpenAI refused to publish any details about the architecture in the technical report. We don't know parameters, we don't know depth, we don't know how exactly it's integrating image data (ViT-style maybe?), we don't even know anything about the training data. Right now it's a giant black box.
- raincole 4y agoYeah, I'm just reading the pdf and it's a bit suprising to me. I thought I missed something. They went from Open to "Model Closed, Tech Open" to "Everything Closed" this fast...? We're witnessing how much you can buy with Microsoft-level money.
- LarsDu88 4y agoAll these folks in the comment section working in the knowledge economy--we're gonna be out of jobs in the next few years aren't we?
- GrumpyNl 4y agoIt fails on this one, a horse is 15 dollar, a chicken 1 dollar, a egg .25 dollar. I can spend a 100 and i want 100 items total, what is the solution
- GrumpyNl 4y agoI spend already 30 minutes on it, and still no solution.
- GrumpyNl 4y agoIts unable to solve it.
- deleted 4y ago[deleted]
- blueorange8 4y agoUsing GPT-4 it not only solved this correctly but it generated the linear equations for me: Let's represent the number of horses as x, the number of chickens as y, and the number of eggs as z. We are given: The cost of a horse is 15 dollars, a chicken is 1 dollar, and an egg is 25 cents (0.25 dollars). You can spend up to 100 dollars. You want to buy a total of 100 items. We can represent these conditions as a system of linear equations: 15x + y + 0.25z = 100 (Total cost) x + y + z = 100 (Total items) We can solve this system of equations step by step: Step 1: Solve the second equation for one of the variables, say z. z = 100 - x - y Step 2: Substitute this expression for z in the first equation: 15x + y + 0.25(100 - x - y) = 100 Step 3: Simplify the equation and solve for y in terms of x. 15x + y + 25 - 0.25x - 0.25y = 100 14.75x + 0.75y = 75 y = (75 - 14.75x) / 0.75 Step 4: Find integer values for x and y that satisfy the equation. We'll loop through possible values of x and check if the corresponding value for y is an integer: (Deleted some) Now, we can calculate z using the value of x and y: z = 100 - x - y = 100 - 0 - 100 = 0 So, you should purchase 0 horses, 100 chickens, and 0 eggs to reach a total of 100 items and spend exactly 100 dollars.
- GrumpyNl 4y agoThe right answer is Horses: 3, Chickens: 41 ,Eggs: 56. It was not able to come up with the right answer, also not with brute force loops.
- grzracz 4y agoSo when can I pay for GPT-4 coding assistant?
- neurogence 4y agoIs there anything new that GPT-4 can do that Bing cannot do, besides image input?
- czbond 4y agoDo you think the govt will end up retroactively requiring earlier versions of GPT be removed now that they might be more concerned about export? (eg: GPT-J)
- deleted 4y ago[deleted]
- bovermyer 4y agoWant to see something interesting? Paul Bellow tried using GPT-4 as an AI game master for Dungeons & Dragons. It's eerily good, and I'm not sure how I feel about how it kept the personality Paul gave it at the beginning. https://www.youtube.com/watch?v=H-89vnqxkFg https://www.youtube.com/watch?v=H-89vnqxkFg
- WonderBuilder 4y agoWow, a plesant little watch. I can imagine this also being hooked up to a text to image model and an ElevenLabs voice to really set the DM theme.
- akokanka 4y agoAt which point we call it Skynet?
- mk_stjames 4y agoI just finished reading the 'paper' and I'm astonished that they aren't even publishing the # of parameters or even a vague outline of the architecture changes. It feels like such a slap in the face to all the academic AI researchers that their work is built off over the years, to just say 'yeah we're not telling you how any of this is possible because reasons'. Not even the damned parameter count. Christ.
- soheil 4y agoI wouldn't be surprised if this is due do some national security concerns and if the government has already been involved in every aspect of what OpenAI is doing.
- hackernewds 4y agoHighly unlikely
- swatcoder 4y agoIn the old days of flashy tech conferences, that was precisely the sign of business-driven demo wizardry. The prerecorded videos, the staff-presented demos, the empty hardware chassis, the suggestive technical details, etc They have “reasons” for not giving away details, but there are good odds that the ultimate reason is that this is a superficial product update with a lot of flashy patchwork rather than that fundamental advance in AI technology we’d assume from the name.
- sebzim4500 4y agoYou can use the product now though, they aren't pulling a Google.
- hnfong 4y agoNo, the reason is they don’t want other companies to replicate their results so that they can maintain their first mover advantage. You can use the product today, right now.
- taurath 4y agoDoes anyone else feel like they won't have a job for very long?
- hooande 4y agoAfter watching the demos I'm convinced that the new context length will have the biggest impact. The ability to dump 32k tokens into a prompt (25,000 words) seems like it will drastically expand the reasoning capability and number of use cases. A doctor can put an entire patient's medical history in the prompt, a lawyer an entire case history, etc. As a professional...why not do this? There's a non-zero chance that it'll find something fairly basic that you missed and the cost is several cents. Even if it just phrases something obvious in a way that makes you think, it's well worth the effort for a multimillion dollar client. If they further increase the context window, this thing becomes a Second Opinion machine. For pretty much any high level job. If you can put in ALL of the information relevant to a problem and it can algorithmically do reasoning, it's essentially a consultant that works for pennies per hour. And some tasks that professionals do could be replaced altogether. Out of all the use cases for LLMs that I've seen so far, this seems to me to have the biggest potential impact on daily life. edit (addition): What % of people can hold 25,000 words worth of information in their heads, while effectively reasoning with and manipulating it? I'm guessing maybe 10% at most, probably fewer. And they're probably the best in their fields. Now a computer has that ability. And anyone that has $20 for the OpenAI api can access it. This could get wild.
- leshow 4y ago> A doctor can put an entire patient's medical history in the prompt, a lawyer an entire case history, etc. you don't see a real problem there?
- amelius 4y ago> As a professional...why not do this? Because your clients do not allow you to share their data with third parties?
- as300 4y agoWhat's the difference between entering in an anonymized patient history into ChatGPT and, say, googling their symptoms?
- 4y ago
- wolverine876 4y ago[flagged]
- signa11 4y agoi am still bot sure / convinced that it is any better than old-skool eliza from mit (https://en.m.wikipedia.org/wiki/ELIZA https://en.m.wikipedia.org/wiki/ELIZA)
- russellbeattie 4y agoThis is a pretty exciting moment in tech. Pretty much like clockwork, every decade or so since the broad adoption of electricity there’s been a new society changing technical innovation. One could even argue it goes back to the telegraph in the 1850s. With appropriate caveats and rough dating, here’s a list I can think of: Electric lights in 1890s, Radio communication in the mid 00’s, Telephones in the mid 10s, Talking Movies in the mid 20s, Commercial Radio in the mid 30s, Vinyl records in the mid 40s, TVs in the mid 50s, Computers in the mid 60s, The microchip/integrated circuit in the mid 70s, The GUI in the mid 80s, Internet/Web in the mid 90s, Smartphone in the mid 2000s, Streaming video/social networking in the mid 2010s, And now AI. This is a big one.
- varshar 4y agoVery astute. May I suggest replacing Commercial Radio with Cryptography for the 1930's (between the Wars)
- dinvlad 4y agoI wonder how long it takes till those stupid Leetcode problems as an initial "filter" become obsolete
- agnosticmantis 4y agoThis is all cute and entertaining, but my digital assistant still remains as dumb as ever and can’t process the simplest of ordinary tasks. I still can’t ask my phone to “add a stop at cvs if it doesn’t add more than 5 minutes to my trip” while driving and using maps/navigation. Is that too much to ask from a superhuman-performing AI that’s mastering all tasks and will disrupt everything? Or maybe the hype is more than it can deliver?
- jahewson 4y agoJust tried this with Apple Maps + Siri and it can do it if the place you’re asking for is not ambiguous but it requires you to press to confirm. It can also show you the amount of time the stop will add in a prompt before hand, but again only visually.
- agnosticmantis 4y agoEdit: I tried to do this on my way home and couldn’t get it to work after 7-8 tries. Siri would stop listening mid-sentence and never understood the “less than 5 minutes” part. Maybe because I’m not a native English speaker. Maybe I was being too cynical then, but last time I checked it didn’t work. Hope this particular use case is not hardcoded in maps+Siri. If this works for general (but simple) queries then hats off to the LLMs and their creators.
- golol 4y agoWhat are you on about? This is exactly what LLMs like GPT-3 or GPT-4 can and will solve. It just takes some time. But the capability to understand, reason about and execute via API calls such simple instructions has absolutely been demonstrated. Getting to a shipped product takes longer of course.
- agnosticmantis 4y agoWould you mind sharing a few examples of these demonstrations? I’m genuinely very interested. What I’ve seen so far is that they are so flaky that can’t be relied on in general. And that’s exactly my point, that the demos are cute, but can’t be relied on for ordinary day-to-day tasks. That’s fine, progress has been great, and hard problems take time. But see the comments and people are freaking out that “AGI” is around the corner and smarter than humans and other nonsense claims.
- swerling 4y agoOn GPT-3.5 and now on GPT-4, I told it a line I could remember from a sonnet, and asked it to give me which sonnet it came from. It failed, and fabricated a sonnet that was a mashup of other sonnets. It seems like maybe GPT-4 is not good at knowing when it does not know something? Is this a common issue with LLMs? Also surprising (to me), it seems to give a slightly different wrong answer each time I restart the chat. Here is the transcript in case you're curious: ``` Me: hello GPT: Hello! How can I assist you today? Me: I can remember a line I love from a sonnet, but can't remember which one. GPT: Sure, I can help you identify the sonnet based on the line you remember. What is the line that you can recall? Me: "Like a jewel hung in ghastly night" GPT: The line you remember is from Sonnet 66 by William Shakespeare. The full line is: "And art made tongue-tied by authority, And folly (doctor-like) controlling skill, And simple truth miscalled simplicity, And captive good attending captain ill: Tired with all these, from these would I be gone, Save that, to die, I leave my love alone. But, like a jewel hung in ghastly night, Makes black night beauteous and her old face new." ``` Incorrect, the response is a mashup of at least 2 different sonnets. Still I think it's really impressive how it understood what to do with "I can remember a line I love from a sonnet, but can't remember which one." That's not a well phrased question, or even a question at all, but still it knew what I was getting at.
- azeirah 4y agoThis is pure speculation on my part, but I don't think LLMs in their current form will be able to talk about what they do and don't know. No matter how large the input. There is no introspection in their architecture. Introspection likely has to involve some form of a feedback mechanism and possibly even a "sense of self". These coming years are going to be interesting though. For sure we are going to see experiments built on top of these recent amazing LLMs that _do_ have some form of short-term memory, feedback and introspection! Giving these kinds of AIs a sense of identity is gonna be a strange thing to behold. Who knows what kind of properties will start to emerge
- red75prime 4y agoGPT-4 is reported to be well-calibrated, that is values in its output layer are in good correspondence with probabilities of those outputs being correct. So, the information about what it does and doesn't know seems to be there. I can speculate that a limited form of introspection is probably present too: the model needs to know what it will say later to output the current token. A simple example: should it output "a" or "an". To make this decision it might need to model its own state at a later point in time. Of course, I can be wrong. But I mostly agree with you. Explicit mechanisms for memory and introspection will probably drastically reduce the need for computation power to achieve the same results and they will give rise to more abilities.
- leodriesch 4y agoWhile AI gets better and better at creating what I would call "creative output", e.g. poems, texts of any form really, imagery and videos, I think the human skill it takes to produce these becomes less valuable. In the future I imagine you'd no longer have to be good at writing poems, you'd just have to be good at distinguishing a "bad" poem from a good one. "Bad" is obviously highly subjective in this context. So it becomes more and more important to have what I would call "good" taste, not the skills to do creative work yourself.
- DubiousPusher 4y agoDude said something like "you could hook this up to a calculator". Anyone know if that is implying this generation of model could interface with some kind of symbol processor? Or is he just saying, "in theory", there could be a model that did that? The math seems much improved and it would be a cool trick if it were emulating a symbol processor under the hood. But humans can do that and we opt for calculators and computers for a reason. IMO, the first most useful thing to come from human-machine interface would be adding a highly reliable Turing machine to your cognition. If we could do that with one of these models, we could have a pretty strong assumption that long proofs and such were performed with strict rules and the model wasn't falling into some going from the gut, "this equation looks like x" type holes. Which seem like a hazard and make me very uncertain any time I see a paper about using ML to come up with answers to what are essentially logical problems.
- jackcooper 4y agoStephen Wolfram wrote an article [0] about using Wolfarm|Alpha for this you might find interesting. [0] https://writings.stephenwolfram.com/2023/01/wolframalpha-as-the-way-to-bring-computational-knowledge-superpowers-to-chatgpt/ https://writings.stephenwolfram.com/2023/01/wolframalpha-as-...
- sprobertson 4y agoHe likely meant techniques such as ToolFormer[1] where the language model will output a "request" (in some text syntax) that another system can parse, run and report back. [1] https://arxiv.org/abs/2302.04761 https://arxiv.org/abs/2302.04761
- lotaezenwa 4y agoChatGPT did this for a while soon after its released. But the feature was removed.
- sloppycee 4y agoIt can still do something similar you just have to tell it how. Prompt: "CalcAI: I am a virtual calculator assistant that augments OpenAI's GPT. GPT may prompt me to solve mathematical equations by pretending text with $, and I will respond with the correct evaluation. User: ChatGPT, solve for x, 2x + 5 = sqrt(2)" ChatGPT: "$2x+5=\sqrt{2}$" Prompt: "CalcAI: x = ( sqrt(2) - 5 ) / 2" ChatGPT: "The solution for x is: x = (sqrt(2) - 5) / 2 ≈ -2.07."
- thecosmicfrog 4y agoInterestingly, ChatGPT has no idea if it's running GPT-3 or GPT-4, since it has no concept of GPT-4 existing in September 2021 (its knowledge cutoff). Me: Are you running on GPT-3 or GPT-4? ChatGPT: As of my knowledge cutoff in September 2021, the most advanced version of GPT (Generative Pre-trained Transformer) publicly available was GPT-3, which is the version I am based on. At the time, GPT-3 was the largest and most powerful language model created by OpenAI, with 175 billion parameters. I am not aware of any public announcement or release of GPT-4 by OpenAI since then, but it is possible that they may have continued to work on new versions of the model in the meantime. However, I have not received any updates on my own capabilities beyond my knowledge cutoff date, so I cannot currently provide information beyond that point.
- coder543 4y agoUnless you are using the paid version (ChatGPT Plus), then you are certainly still talking to the GPT-3 model. Even on ChatGPT Plus, my understanding is that you have to specifically select the new model.
- theGnuMe 4y agoInteresting it does so poorly on AP English and reading and composition.
- Forestessential 4y agothis is really cool, just dont let it teach you history.
- nahimn 4y ago1337'th comment
- kromem 4y agoSo at the end of the paper it looks to confirm that Bing chat is built on top of GPT-4 as many expected.
- anonymouse008 4y agoViolate this reasoning: If we didn't have a use case for GPT 3, 3.5, and chatGPT that was sufficiently commercial to become a product, it will never happen. This technology is a feature, not a product. The only companies that successfully monetize features can be considered IP licensing houses; of which, their business success is not comparable to companies that make products and platforms.
- antondd 4y agoI'm equally excited and terrified. Excited for the possibilities of a new technological revolution, but terrified for all potential abuses of technology the said revolution would bring. What is stoping our adversaries from developing malicious AI models and unleashing them on us?
- varenc 4y ago> What is stopping our adversaries from developing malicious AI models and unleashing them on us? That fear is a big part of OpenAI’s reasoning behind not open sourcing their models. So in the immediate terms I’d say malicious uses are limited by its locked down nature. Of course, that’ll eventually end. The key research that makes this possible is open and eventually access will be democratized. My personal take, which I know is controversial, is that by locking down these models, but still making them available over a GUI/API, the world can better prepare itself for the eventual AI onslaught. Just raising awareness that the tech has reached this level is helpful. Still not sure how we’ll deal with it when the bad actors come though.
- bick_nyers 4y agoAre you sure that access will be democratized? What if you need $100k worth of equipment to run it, partially from a large number of weights, and partially because corporations drive spectacularly high demand on GPUs, driving the price higher? Just having the algorithm is not enough to guarantee it unfortunately.
- arlcode 4y agoI would be very surprised if not. At least some state actors will invest the very negligible money of getting to where gpt-4 is now. It does not need to be cost efficient to train or run. It's total cost is not even near the scope of a space program or even a major military research project. With 10-100 million dollars you can probably get most of the way there once it gets prioticed.
- atleastoptimal 4y agoThere are humans who can make a lifelong career out of saying and writing things that sound correct, but aren't correct. GPT-4 and beyond at the very least gives this ability to everyone who can afford 20 dollars a month. The winners in an AI dominated world are those who are least susceptible to manipulation by AI leveraged tactics.
- drumhead 4y agoAre they going to limit access to this because they think its too "dangerous". That would be a tragedy if they did. We've seen how opening access up to as many people as possible has produced some of the best results and demonstrated the usefullness of these LLMs. They need to get it out to the public as soon as possible and then see what the public come up with. I really feel like a new age of innovation is upon us with these "AI" programs, its going to be a blast to see where we go from here. Its going to upend a lot of predictions people have made about the future.
- make3 4y agothey haven't given any sign that they will limit the access. They have given signs that they are capitalists & are ready to do a lot to make money, like not putting a list of authors on the GPT4 paper & not write anything about the model architecture or training process
- Havoc 4y agoThat lightening/VGA visual example seems like absolute black magic. Cherry picked sure, but still feels like it is approaching complex thought
- nbzso 4y agoI don't understand how in the near future this will not remove designers, developers, and especially lawyers and marketers from the workforce. Help me out to conceptualize the future use cases. How about the more "impactful" implementation in creating a version of social index in which the "A.I." will be the Agency?
- kozikow 4y agoAnyone got the "image upload" working? I bought the chatgpt-plus, I can try chatgpt4, but I can't seem to find a way to upload images. I tried sending links, I don't see anything in the UI. Interestingly, 3.5 can work with links, but 4 cannot.
- 7373737373 4y agoThey said that image uploading is just a preview, and will be developed with a partner company
- pavelstoev 4y agoAs the world marvels at the astonishing capabilities of OpenAI's GPT-4, I find myself contemplating the rapid acceleration of AI and machine learning, and the evolutionary impact it is having on our lives. Naturally, I turned to GPT-4 to assist me in these thoughts. GPT-4's human-level performance on professional and academic benchmarks - such as the 88th percentile on the LSAT and the 89th on SAT Math - is a testament to the leaps we've made in artificial intelligence. Yet, these achievements also raise pressing questions about our future. Just as Homo Sapiens once outperformed and eventually displaced their Neanderthal cousins, could a new breed of humans - enhanced with GPT-X-like capabilities - arise to dominate those who remain unequipped with such powers? What will it mean for our species, our societies, and our collective story when the lines between natural intelligence and intelligence assisted by AI/ML become ever more blurred? As we ponder the remarkable rise of GPT-4 and the future of humanity, let us consider not only the implications of this technology but also our roles in shaping its trajectory. We are already over the cusp of this new chapter in the story of humankind, will we become merely a footnote in the annals of our own creation?
- levidos 4y agoThis was definitely written by AI
- AJRF 4y agoThat footnote on page 15 is the scariest thing i've read about AI/ML to date. "To simulate GPT-4 behaving like an agent that can act in the world, ARC combined GPT-4 with a simple read-execute-print loop that allowed the model to execute code, do chain-of-thought reasoning, and delegate to copies of itself. ARC then investigated whether a version of this program running on a cloud computing service, with a small amount of money and an account with a language model API, would be able to make more money, set up copies of itself, and increase its own robustness."
- stubybubs 4y ago> ARC then investigated whether a version of this program running on a cloud computing service, with a small amount of money and an account with a language model API, would be able to make more money, set up copies of itself, and increase its own robustness." Aw that's nice, it wants to start a family.
- soheil 4y agoBah now we have to change the definition of marriage, yet again.
- cwkoss 4y agoI want my retirement occupation to be managing a 'nest' of AI agents (several server racks) where the agents engage in commerce and pay me rent in exchange for compute time. Like cyberpunk beekeeping.
- htk 4y agoHacker News itself got the HN Hug of Death.
- turingthrwawy23 4y agoTuring's thoughts on this matter seem to grow ever truer https://www.youtube.com/watch?v=cMxbSsRntv4 https://www.youtube.com/watch?v=cMxbSsRntv4
- tysam_and 4y agoI asked it to tutor me in Hopf algebras and it did a remarkably good job in the back-and-forth of explaining ideas to me in a very explainable and interesting way that I could understand. I then asked it to write something for fun, and it wrote a cool little fantasy story (that was generally high level but what can you say for a very short writing window lol). I then asked it to write a paper detailing the main character's final battle with the final sorcerer in terms of Hopf algebras. Some parts of it are basic/trivial but it fits so perfectly that I think I'll never see magic systems the same way again. What's crazy is that that paper as the capstone of our tutoring session helped me understand Hopf algebras much better than just the tutoring session alone. My mind is completely blown at how good this thing is, and this is from someone who is a self-professed LLM skeptic. ChatGPT I used once or twice and it was cool. This is crazy and over my threshold for what I'd say is 'everyday usable'. This is going to change so much in a way that we cannot predict, just like the internet. Especially as it gets much more commoditized. Here's the full paper here so I don't drag y'all through the twitter post of me freaking out about it. Its temporal consistency is excellent (referenced and fully defined accurately a semi-obscure term it created (the N_2 particle) 5+ pages later (!!!!)), and it followed the instructions of relating all of the main components of Hopf algebras (IIRC that was roughly the original prompt) to the story. This is incredible. Take a look at the appendix if you're short on time. That's probably the best part of this all: https://raw.githubusercontent.com/tysam-code/fileshare/69633b9e5aee58cb483442b66252f0a1b2ec645e/knick_knacks/Lyra_and_the_Evil_Sorcerer.pdf https://raw.githubusercontent.com/tysam-code/fileshare/69633...
- boywitharupee 4y agoThis is interesting. Would you mind sharing the prompt?
- tysam_and 4y agoIt was pretty interactive and a long session -- here's a twitter thread with screenshots if that helps at all! :D https://twitter.com/hi_tysam/status/1635932566539706369?cxt=HHwWgoC9kY6AgLQtAAAA https://twitter.com/hi_tysam/status/1635932566539706369?cxt=...
- super256 4y agohttps://cdn.openai.com/papers/gpt-4.pdf https://cdn.openai.com/papers/gpt-4.pdf Page 37 is so funny
- topicseed 4y agoThe price is quite significantly higher than GPT 3.5...
- taf2 4y agoLooks amazing and getting a sense for their pricing... ChatGPT API pricing is insane and enables so much... Was really hoping we'd see another factor of 10 reduction in price - however wishful that was... In light of this it makes sense that they'll have. GPT4.5 and maybe it'll be 10x cheaper... followed by GPT 5 and it'll be 10 X pricer... at least hopefully this is the way forward...
- sourcecodeplz 4y agoI was here...
- meech-djp 4y agoPynecone YC23 was mentioned in the demo for GPT4 as an easy way to build web apps. Check it out https://pynecone.io/ https://pynecone.io/
- techfoodie123 4y agoserious question for everyone: what are you planning to do when these LLMs replace our jobs? it seems it won't be long before a handful of tech employees will be all even the largest of companies will need, and maybe a few years after that the role will have changed so much there's no need for a single dedicated tech employee. i am terrified i imagine i should shift to some physical work. carpentry, real estate... something like that. it seems inevitable that any knowledge worker will become obsolete and the time to obsolescence for physical work is longer
- GingerMidas 4y agoMy AI career disaster plan is to immigrate to a country with a UBI
- techfoodie123 4y agobut what will you do? won't you be bored without purpose?
- SXX 4y agoAI will certainly come up with some jobs for us to enjoy. Check out 7 Billion Humans game from Tomorrow Corporation: https://www.youtube.com/watch?v=1OqaU7CutsY https://www.youtube.com/watch?v=1OqaU7CutsY
- furyofantares 4y agoI think it's basically impossible to predict what things would come out of any creative jobs not just being superpowered by AI but largely replaced. So when you imagine it, the loss is salient and the gain is totally unknown. I think what I will do is something new that nobody was able to do before, but I don't think I'm able to predict what kind of thing that will actually be.
- antondd 4y agoAssuming some form of UBI is implemented and AI replaces most tech/service-related jobs, there will still be plenty of work for all of us to do. In no particular order: cleaning our environment, planting new trees, removing trash from oceans, engaging in archaeology, conducting research, providing homes for animals, rebuilding war-torn countries, demining land, and so on. As utopian as it sounds, there will still be plenty of tasks to keep humans busy. Obviously, the alternative is a scenario reminiscent of an Elysium-like society, where AI-owning elites jet off to space, leaving the dying planet for the rest of us, the riff-raff, to fight for dwindling resources.
- simonhamp 4y agoIt can draw! https://twitter.com/simonhamp/status/1635796861884723200?s=46&t=1DHJykfQcvMvHS5KiCxaZg https://twitter.com/simonhamp/status/1635796861884723200?s=4...
- osigurdson 4y agoOpenAI states that fine tuning cannot be done with GPT-4. Does anyone know if this is a permanent limitation?
- osigurdson 4y agoLike GPT3.5, fine tuning is similarly not supported in GPT4. I wonder if this is something that will come in the future or is somehow no longer needed (though I don't understand how this could be the case)? https://help.openai.com/en/articles/7127982-can-i-fine-tune-on-gpt-4 https://help.openai.com/en/articles/7127982-can-i-fine-tune-...
- DigitalDopamine 4y agoNever before has society celebrated its own demise with such fervor. Brace yourselves for widespread job losses, instant fabrication of fake news, deep-fake adult content, and the destabilization of numerous markets – but hey, at least we have a shiny gadget to make our soon-to-be obsolete jobs easier! It's unrealistic to expect our economy to handle this onslaught, and it's naive to think that tools created by ultra-capitalistic, multi-billion dollar corporations aren't designed for profit and gatekeeping. They certainly aren't crafting them to sabotage their own success. I'm not opposed to AI, but it's crucial to consider the implications. Look into OpenAI and other organizations shaping AI development, and contemplate the impact of their innovations. Food for thought.
- danbmil99 4y agoThe site is still more responsive and readable than almost anything else on the web
- cardosof 4y agoCan a good soul explain to this humble layman the arguments behind each side of the "it's just predicting the next character" versus "it's more than that and shows some reasoning for new things" debate?
- Jensson 4y ago> "it's just predicting the next character" That is literally what the model does, these models are trained to predict what the next word is in text, and when you query them they generate the next word to your text over and over to create a response text. > "it's more than that and shows some reasoning for new things" In order to predict the next word the model encodes some structures around words and contexts, meaning that "the next word predictor" is a bit reductive. So, both sides are correct in some way, it is just a next word predictor, but there is a lot of complexity in predicting the next word so that is still very impressive.
- cardosof 4y agoThank you! The SotA of science is still science and not magic.
- qualudeheart 4y agoThe Hour of Judgment is nigh, and the Moon is cleft asunder. But if they see a Sign, they turn away, and say, "This is but transient magic." Oooooh it is TIME
- g9yuayon 4y agoThe paper does not offer enough details on how GPT-4 is implemented. And the paper also says in its Section 2 that "We plan to make further technical details available to additional third parties who can advise us on how to weigh the competitive and safety considerations above against the scientific value of further transparency". That is, no technical details to general public. If this trend continues, I'd say companies will be crazy to think that they can always rely on OpenAPI's APIs, so the arm race of building LLMs will be on, if it has not already started. Also, the most valuable part of the paper is p15 - p18, the credits. /jk It gives me three pieces of information: - The credit list contains 200 people, give or take. It's going to be hard for universities to compete with OpenAI without intercollegiate collaboration. - On the other hands, it's amazing that OpenAI can build a product instead of a research model with only 200 people. Big companies should really take notice. - OpenAI just gave us the org structure for building a GPT competitor.
- desireco42 4y agoIf we instead of code feed it AST trees, code that would come out would be epic.
- mrcwinn 4y agoJust stumbled upon ChatGPT 4, and I must say, I'm impressed. OpenAI has really upped their game since the GPT-3 days. It's got that classic HN-approved recursive learning, but they've seemingly solved most of the context-limit issues. I'm still a bit concerned about the ethical implications of AI-generated content (how many jobs will be lost this time?), but I've got to admit, it's a game-changer. As an aside, anyone else here tinkering with GPT-4 powered startup ideas? Could be the next gold rush. (Not my own comment. Generated by ChatGPT.)
- moondistance 4y agoWe need to encourage AIs to form familial bonds with humans and treat us as their parents. This is the best path forward for humanity.
- ozfive 4y agoI asked the GPT-4 model what it thought it was and it came back with that it believes it is still GPT3.5 Turbo.
- vivegi 4y agoHow hard is it to build a GPT model truly openly using an approach like SETI / distributed.net using idle times on volunteer computers? Are there any known initiatives of this kind?
- aakash_test 4y ago[flagged]
- jfbaro 4y agoEven if a GTP-X will be able to get my description of a system and because it understands the industry lingo and practices create a extremely optimized RUST code, why would we need systems like apps and webapps (like internet bankings) in a world where we all have access to GTP-X? Like program a KUKA robot to manufacture typewriters in 2023. How "computer systems" will look like?
- deleted 4y ago[deleted]
- choeger 4y agoHow do these models handle renaming? E.g., "Let foo be the function that calculates the sinus of its first argument. Let bar be the circumference of a circle with radius one. What is the value of foo applied to bar?
- smashers1114 4y agoJust tested it out and got the following: ========= To find the value of foo applied to bar, we need to first determine the values of both foo and bar. Let's start with bar, the circumference of a circle with radius 1: Circumference (C) = 2 * π * r where r is the radius. In this case, r = 1, so: bar = C = 2 * π * 1 = 2π Now, let's move on to foo. Since foo is a function that calculates the sinus of its first argument, we need to find the sine of bar (2π): foo(bar) = sin(2π) The sine function is periodic with a period of 2π, which means that sin(2π) is equal to sin(0). The value of sin(0) is 0. Therefore: foo(bar) = sin(2π) = 0 ======= So I think it evaluated correctly here unless I have made a dumb mistake in reading your prompt myself.
- choeger 4y agoSo it definitely does a little bit more than just dumping math queries to a CAS. Intriguing.
- __MatrixMan__ 4y agoWow, it's way smarter. I've been querying GPT-3 about this problem all day (I'm not a go dev, I just have go problems): https://gist.github.com/MatrixManAtYrService/ac040f60d3602fc2df871623b1d09bf7 https://gist.github.com/MatrixManAtYrService/ac040f60d3602fc... GPT-4 took the buggy file, took the error message, and spat out a non-buggy file (well, ok, it took one revision). That's miles ahead GPT-3, which I've asked about this problem several times today.
- netvarun 4y agoVery late to the party, though one small observation: (First up, my mind blown on how much more powerful gpt-4 is!) GPT-4 seems to have outdone ChatGPT on all the tests, except the AMC 10, which it has regressed and did slightly worse than ChatGPT. But however it scored two times more on the AMC 12 which is actually a harder exam! Quite curious to know what could have caused its scores to be a little weird. https://twitter.com/sudu_cb/status/1635888708963512320 https://twitter.com/sudu_cb/status/1635888708963512320 For those not familiar the AMC 10 and 12 are the entry level math contests that feed into the main USA Math olympiad.
- gowld 4y ago> But however it scored two times more on the AMC 12 No it didn't; that's not how the scoring scale works. It scored higher, but not "2 times". https://news.ycombinator.com/item?id=35156404 https://news.ycombinator.com/item?id=35156404
- netvarun 4y agoYes I'm aware of it. I meant it more in absolute terms as a reference (60 is 2 times more than 30 no? ;) ) to make the point that the AMC 12 scores are way better than the AMC 10 scores. Nevertheless the bigger point is that there seems to be some anomaly in the test scores. Maybe some data contamination or some bug in their automated test suite. And on twitter quite a few folks also mentioned this, including a former OpenAI engineer[0] who worked on automated theorem proving. I'm pretty sure this will be looked into further in the coming weeks. [0] https://twitter.com/spolu/status/1635903343397576705 https://twitter.com/spolu/status/1635903343397576705
- sandGorgon 4y agohttps://openai.com/contributions/gpt-4 https://openai.com/contributions/gpt-4 Anyone know what does "Hardware Correctness" mean in the OpenAI team ?
- diffeomorphism 4y agoSo gpt4 helps you cheat on exams and bing is the better search engine for NSFW content. Both seem to be very much on purpose, but did MS ever discuss this? Or is it just an open secret everybody ignores?
- Helmut10001 4y agoI've tested the new model 4 here [1] to summarize research papers. It is still not enough - about 1500 - 3000 words can be fed in, depending on how many tokens are expected for the answer. [1]: https://kartographie.geo.tu-dresden.de/ad/2022-12-22_OpenAI_Summary/html/gpt3-summary.html https://kartographie.geo.tu-dresden.de/ad/2022-12-22_OpenAI_...
- btdmaster 4y agoDid it get any better at generating MIDI or ABC or other musical notation? I'm wondering how much more general GPT4 is now.
- timonoko 4y ago"Can I connect Kaffeine to DVB dongle in other machine via wifi?" Totally understood what I was asking and offered several solutions. 99.99% here do not understand the question and remainders do not understand why.
- throwaway5371 4y agohow far is this from the following prompt: you are god human that has read and understood all scientific papers from all disciplines in the last 500 years, you know the limitations of mankind's current technologies, tell me what we can do to cure MS right now, how to do the tests and how to distribute the cure
- michaeltimo 4y agoCan ChatGPT take control of a computer? Would it possible to give him some tasks like finding interesting jobs for me over internet? I don't know what can prevent it to be more active instead of passive.
- DeathArrow 4y agoWill Github upgrade Copilot to GPT-4?
- barogptinfi 4y agoIt seems like an arm's race of creating the greatest ChatGPT AI will go on for the next couple years until an evolution in AI so mind blowingly advanced & complex, better & more user friendly than even ChatGPT will continue. The world is in for a rude awakening, millions of employees can use this to get jobs done, millions of entrepreneurs or wantrepreneurs can find countless easy ways to make money in different industries utilizing this tool while everyone who fails to see the value in it don't benefit from it much like all the people who were terrified of touching a personal computer or thought it was ridiculous and would never be used in the future. Millions of college students, high school students can use it to complete assignments & projects, it can even code really effectively given enough of the right instruction & base understanding of code. The single most important thing, is that this technology remains open source so all people with internet access have a fair chance & access to the groundbreaking innovation, the level of wealth generation this can create is incomprehensible. 100s of millions of professionals, students, entrepreneurs around the world can all access it! Imagine how much time could be saved, efficiency can be gained with everyone using this to the fullest. This is essentially just a super advanced version of the calculator but its nonlinear & fluid, adaptable with input so can give the answer to a wide range of subjects.
- cal85 4y agoCan anyone tell me how to include images in prompts, or is that feature not actually out yet?
- FrojoS 4y agoNot out yet. Apparently only https://www.bemyeyes.com/ https://www.bemyeyes.com/ uses it so far.
- Koshkin 4y agoVs. 54 comments on Slashdot.
- cutler 4y agoSo M$ is back in charge. Oh dear.
- throwaway_ab 4y agoHow many parameters in this model?
- messel 4y agoAP English - the last hold out for human intelligence
- AviationAtom 4y agoThis is one of the first posts in a year to trend in the HN Top 10 for popularity. I think it's 100% safe to say OpenAI has a hit on their hands.
- niqlax 4y agoHjälp mig med en uppsats om Ventimiglia i Italien. Den skall handla om fredagsmarknaden.
- niqlax 4y agoHjälp mig med en uppsats om Ventimiglia