2 ms·
Sort of just proves my point, no? It's faster for me to just check its work than to do the work from scratch myself and that work isn't monopolizing the wetware
by barking_biscuit 4y ago
Sort of just proves my point, no? It's faster for me to just check its work than to do the work from scratch myself and that work isn't monopolizing the wetware of another human being. Cognitive bandwidth has increased. The system is capable of more throughput than before. In fact, how much more throughput could you get if you employ enough instances of LLMs to saturate the cognitive bandwidth of a human who's sole job is to simply verify the outputs?
If you look at the leap from GPT-3 -> GPT-4 in terms of hallucinations and capabilities etc, and you combine that with advances in cognitive architecture like reflexion and AutoGPT, it's pretty clear the trajectory is one of becoming increasingly competent and trustworthy.
The degree to which you need to check it's work depends on your use case, level of risk tolerance, and methods for verification. I think one of the reasons AI art has absolutely exploded is because there's no consequences for a generation that fails and it can be verified instantly. Compare that to doing your taxes where it's high stakes if you get it wrong, you're far less likely to rely on it. There is a landscape of usefulness with different peaks and valleys.
- ChatGTP 4y agoWhat are one of the professional use cases where you would just feel comfortable YOLOing some ChatGPT generated code into prod? Publishing a journal without verification etc? You should also take note of the warnings in the GPT-4 manual, it's a much more convincing liar than GPT-3. Quite explicitly says that. My fear is that I just get lazy and trust it all the time. I think one of the reasons AI art has absolutely exploded is because there's no consequences for a generation that fails and it can be verified instantly. What are you talking about exactly?
- barking_biscuit 4y agoWhat's with the assumption that anyone needs to YOLO anything? Your coworkers don't let you YOLO your code to prod, and you don't let them YOLO their code to prod. Trust but verify, right? My point with the AI art comment is that not every output of these models is something that needs to go to production! There's a continuum of how much something matters if it's wrong, and it depends on who is consuming the output and what it is they need to do with it, and the degree to which other stakeholders are involved.