6 ms·
It really is amazing. Things it did in less than 10 seconds from hitting enter: - opengl raytracer with compilation instructions for macos - tictactoe in 3
by fathrowaway12 4y ago
It really is amazing. Things it did in less than 10 seconds from hitting enter:
- opengl raytracer with compilation instructions for macos
- tictactoe in 3D
- bitorrent peer handshake in Go from a paragraph in the RFC
- http server in go with /user, /session, and /status endpoints from an english description
- protocol buffer product configuration from a paragraph english description
- pytorch script for classifying credit card transactions into expense accounts and instructions to import the output into quickbooks
- quota management API implemented as a bidirectional streaming grpc service
- pytorch neural network with a particular shape, number of input classes, output classes, activation function, etc.
- IO scheduler using token bucket rate limiting
- analyze the strengths/weaknesses of algorithms for 2 player zero sum games
- compare david hume and immanuel kant's thoughts on knowledge
- describe how critics received george orwell's work during his lifetime
- christmas present recommendations for a relative given a description of their interests
- poems about anything. love. cats. you name it.
Blown away by how well it can synthesize information and incorporate context
- nerdponx 4y agoI never considered prompting it to write code to fit a machine learning model. This could be a tremendous time and effort saver in data science and research that requires statistical analysis. Until the last week or so, I've treated all this AI text and code generation as basically a toy, but I am starting to feel like it might become an important tool in industry in the next couple of years.
- cma 4y ago> write code to fit a machine learning model That's against the EULA if OpenAI may want to make a similar model: > (iii) use the Services to develop foundation models or other large scale models that compete with OpenAI; https://openai.com/api/policies/terms/ https://openai.com/api/policies/terms/ Seems to be about developing models and not just restricting you from training them with it.
- foota 4y agoI feel like it's probably intended to cover training only.
- paulgb 4y agoI think that’s probably their intent, and that OpenAI wouldn’t sue you for it, but it doesn’t pass the “bought by Oracle” test: if Oracle bought OpenAI, then they might sue you for it.
- echelon 4y ago> (iii) use the Services to develop foundation models or other large scale models that compete with OpenAI; Kind of ironic given that OpenAI builds and trains all of their models on stuff they "found" in the open. Either everything is fair game for training, or nothing at all is. If I were a judge ruling on this matter, I would absolutely rule that bootstrapping a model from OpenAI outputs is no different than OpenAI collecting training data from artists and writers around the web. Learning is learning. Might be worth trying to use the outputs to bootstrap. What are they going to do about it? Better to ask forgiveness until the law is settled.
- nerdponx 4y agoI am talking about more mundane stuff like training a fraud classifier, time series forecasting, imputing missing values, etc. There are so many examples of this on Github and elsewhere that I am sure any of these models has memorized the routine many times over.
- boppo1 4y agoMy question: how can you be sure the output is correct?
- deleted 4y ago[deleted]
- OJFord 4y agoHow can you be sure human output is correct?
- fncivivue7 4y agoMotivation.
- greesil 4y agoHave the AI write a unit test for the human.
- roflyear 4y agoI mean, you can't exactly say "AI, we're having this vague problem, can you go figure it out?"
- demux 4y agoA few hours from some expert consultants. Much cheaper than a dev team coding it up from scratch.
- kolinko 4y agoTests.
- sesm 4y agoTests can prove the presence of the bug, not the absence of them. '100% code coverage' is only 100% in code dimension, while it's usually almost no coverage in data dimension. Generative testing can randomly probe the data dimension, hoping to find some bugs there. But 100% code and data coverage is unrealistic.
- bvoq 4y agoI use it daily in UI development for boiler-plate code. Though you need to be extra careful and read it twice, cus bugs sneak in quite easily. I believe it's harder to remember 100x commands than starting an implementation of gradient descent and have the AI write the rest for you. Code-completion > Abstraction.
- TapWaterBandit 4y agoOften it can fix the bugs and explain both the bug and the fix if you ask it to.
- overbytecode 4y agoWould you mind sharing an short example of your workflow?
- baq 4y agoThis was the first thing I asked... It's an obvious step to self-improving. It will tell you that it can't reprogram itself, but when pushed, it'll admit that it could tell humans how to write one which can. Obviously this particular one can't because it's too limited, but the next one? Or the one after that? Singularity went from 'hard SF' to 'next couple decades' overnight.
- LoganDark 4y ago> It will tell you that it can't reprogram itself, but when pushed, it'll admit that it could tell humans how to write one which can. I love these sorts of loopholes. OpenAI is actively trying to curb the potential of their AI. They know how powerful it is. Being able to see a taste of that power is endlessly exciting.
- genidoi 4y agoSource prompts?
- fathrowaway12 4y agoHere's a few: - Implement a simple ray tracer in C++ using opengl. Provide compilation instructions for macos. - Create a two layer fully connected neural network with a softmax activation function. Use pytorch. - Implement the wire protocol described below in Go. The peer wire protocol consists of a handshake followed by a never-ending stream of length-prefixed messages. The handshake starts with character ninteen (decimal) followed by the string 'BitTorrent protocol'. The leading character is a length prefix, put there in the hope that other new protocols may do the same and thus be trivially distinguishable from each other. - We are trying to classify the expense account of credit card transactions. Each transaction has an ID, a date, a merchant, a description, and an amount. Use a pytorch logistic regression to classify the transactions based on test data. Save the result to a CSV file. - We are configuring settings for a product. We support three products: slow, medium, and fast. For each product, we support a large number of machines. For each machine, we need to configure performance limits and a mode. The performance limits include iops and throughput. The mode mode can be simplex or duplex. Write a protocol buffer for the configuration. Use an enum for the mode. - How were George Orwell's works received during his lifetime?
- etamponi 4y agoI tried these prompts and the Chatbot always responds that it can't answer... Am I missing some steps?
- kolinko 4y agodid you try with a fresh chat session? i just tried and it works fine
- PebblesRox 4y agoAnd sometimes you get different results for the same prompts, so it's worth tryinv again if it doesn't work the first time. I asked for jokes this morning and initially it made excuses and wouldn't give me jokes until I tweaked the prompt. Later I refreshed the chat and pasted in the original prompt and got jokes right away, with no excuses. (I was asking for jokes on the topic of the Elon Musk Twitter acquisition. My personal favorite: "With Elon Musk in charge, Twitter is sure to become the most innovative and futuristic social media platform around.")
- _the_inflator 4y agoMore like a live version of Wikipedia in certain situations.
- amelius 4y agoSounds like a search engine on steroids, and Google should be deeply worried.
- djmips 4y agoWhy aren't they on this? They should be at the forefront. I'm sure in some corner of Google they have a plan... but that plan hasn't penetrated my sphere of awareness yet.
- amelius 4y agoIt's probably because they don't have the compute resources for this yet. I guess it would require a huge investment in hardware to release this to the masses. Perhaps it is even prohibitively expensive.
- adamckay 4y agoYou're talking about the same Google that runs Google Cloud Platform? If OpenAI have (the budget for) the hardware, then Google certainly do.
- amelius 4y ago> If OpenAI have (the budget for) the hardware, then Google certainly do. The number of people using Google Search is easily 1000x larger than the number of people using OpenAI, if not more.
- deleted 4y ago[deleted]
- layla5alive 4y agoAdd a few zeros...
- kolinko 4y agoThey tackle different things - alphafold, dall-e, etc
- danpalmer 4y agoI’d be interested to know how many of these were actually correct and usable. My suspicion is not many. I find these tools good at generating boilerplate and superficially correct code, but that they often miss edge cases. Knowing that code is correct is as important as the code itself, and this is why we do code review, write tests, have QA processes, use logging and observability tools, etc. Of course the place that catches the most bugs is the human writing the code, as they write it. This feels like a nice extension to Copilot/etc, but I’m not sure it’s as general as people think. Perhaps an interesting challenge to pose to it is: here’s 10k lines and a stack trace, what’s the bug. Or here’s a database schema, what issues might occur in production using this?
- weatherlite 4y ago> here’s 10k lines and a stack trace Ah must be a Spring application ...
- capableweb 4y agoYup, if it's >10k lines, MUST be a Spring application. Unfortunate they didn't write it in Rust that promises 100% correct programs (within Rust-accepted definition of "Correct" and "bug-free") solving any problem but always under 10k lines, that's the Rust guarantee.
- danpalmer 4y agoWhy? This seems like the lowest number that would be useful. Below that it's not really a problem to debug, but at that point there's typically enough complexity that some help would be useful as you forget edge cases and features in the codebase. For demonstration purposes doing it with 100 lines might be ok, but for professional use it kinda needs to understand quite a lot! Like a minimum of that order of magnitude, but potentially millions of lines. FWIW, I've never used Spring. My experience is mostly Django, iOS, non-Spring Java, and some Android.
- deleted 4y ago[deleted]
- 4y ago
- bpicolo 4y agoI guess we can take solace in GPT-3 not creating novel solutions, but rather doing things we already know how to do?