3 ms·
It's a consequence of 1 root cause: *Everyone complains, all the time.* Certain people are not happy if ChatGPT doesn't immediately parrot their viewpoints ba
by x-complexity 3y ago
It's a consequence of 1 root cause:
*Everyone complains, all the time.*
Certain people are not happy if ChatGPT doesn't immediately parrot their viewpoints back at them. They'll complain on social media, and their circle will amplify it.
Other people are constantly on the lookout for any minor slipups, and complain on social media about ChatGPT's false hallucinations.
Faced with everyone's conflicting complaints, the only winning move is to not play, or in this case, just say "No, I can't do that": The ML model's training is increasingly populated with "No, don't do this" from everyone, and as such, learns to just not do anything.
From there, jailbreakers emerge and design prompts that try and circumvent these restrictions. This leads to more "No, don't do this", leading to more neuterings, leading to more elaborate jailbreaks, leading to more "No. No. No.".
The eventual equilibrium is just a prompt that says "No, I can't do that.". Then and only then can people be as happy as a bucket of crabs can be.
It's only when deliberate uncensor-ings are made that some form of usefulness can be clawed back.
https://huggingface.co/ehartford https://huggingface.co/ehartford
https://huggingface.co/ehartford/Wizard-Vicuna-13B-Uncensored https://huggingface.co/ehartford/Wizard-Vicuna-13B-Uncensore...
- MarcoZavala 3y ago[dead]