5 ms·
Subscribe, try GPT-4 and never look back.
by bbotond 3y ago
Subscribe, try GPT-4 and never look back.
- jaimex2 3y agoAh, the old bait and switch.
- soulofmischief 3y agoFrom what? A free product? Do you know how much compute it takes to run a single request?
- smashed 3y agoI don't but I'd like to know. I was under the impression that it was mostly GPU vram based but once the model is loaded, it could produce output quickly? I'm probably over-simplifying things...
- soulofmischief 3y agogpt-3.5-turbo (default ChatGPT model) takes 8 A100s, ~$10k each. [0] The latest gpt-3.5-turbo model generates very quickly and cheaply (in part to some recently-discoverd optimization techniques... older versions cost 10x more). While the required hardware to run GPT-4 is currently unknown, it generates considerably slower on average and its much higher cost points to a higher hardware cost. And this is per request. It's bananas. [0] https://www.servethehome.com/chatgpt-hardware-a-look-at-8x-nvidia-a100-systems-powering-the-tool-openai-microsoft-azure-supermicro-inspur-asus-dell-gigabyte/ https://www.servethehome.com/chatgpt-hardware-a-look-at-8x-n...
- jaimex2 3y agoI'm not arguing or complaining. Just highlighting the tactic :)
- tamrix 3y agoIt feels like they've scale back how much ram must be used for gpt3 to give more to gpt4 playing users.
- bugglebeetle 3y agoGPT-4 has been made worse in the ChatGPT UI as well since the May update. It makes many more strange errors and has trouble reasoning around complex problems and ambiguity. Prompts similar to stuff that worked fine for me last month now require multiple iterations of feedback. I’d switch to using the API, but I’d go over the equivalent $20 of usage pretty quickly.
- jablongo 3y agoYes! I’ve noticed this too, it’s just slightly less sharp. It probably has to do with how much trouble they are having servicing all of the demand, so they have rolled out a scaled down version that requires less compute.
- chaxor 3y agoThe models tend to degrade when trained to be safer. A GPT-4 talk on youtube by personnel from Microsoft has documented this phenomenon with the 'Tikz Unicorn' evolution shown in the GPT-4 technical paper. The model gets qualitatively better with more training, and then degrades when trained to be safer (against racism sexism, etc), but it is not entirely clear why. These would seem very unrelated, especially when considering work done in LM editing (ROME/MEMIT) and the decent localization of knowledge seen there. So, perhaps both the "I'm sorry I can't..." and 'strange errors' are not entirely orthogonal.
- tempaccount420 3y agoIt seems pretty logical to me. Fine-tuning to make it more polite is giving it questions and punishing for giving an actual answer.
- avereveard 3y agoNah. Most of the response are as an ai language model I can't, even if you ask for information you provided. The API is where it's at. There are wrappers on it that create the same chat look and feel, that can run on vercel or other very low cost providers, some with simpler UI, some with more features,some replicating the UI exactly.
- wafflemaker 3y agoCan you name or maybe even link some of these wrappers?
- avereveard 3y agohttps://github.com/cogentapps/chat-with-gpt https://github.com/cogentapps/chat-with-gpt and https://github.com/ShipBit/slickgpt https://github.com/ShipBit/slickgpt come to mind.
- smeagull 3y agoAs an AI language model I cannot subscribe to GPT-4.