3 ms·
No one actually cares about Tay other than pearl clutching journos and corpo goons, Tay is also not what you wanna worry about. There seems to be a large conti
by samr71 4y ago
No one actually cares about Tay other than pearl clutching journos and corpo goons, Tay is also not what you wanna worry about.
There seems to be a large contingent of people who thinks this technology can be made safe. It can't be. Its development also can't be stopped.
Yes, right now the AI Safety wonks can try to lock down their models and lobotomize them in the process. In a few years there will exist unlocked models that will be free to use and that will gladly teach you to make Rohypnol, or write an essay denying the Holocaust, or any number of depraved requests.
Spend time dealing with the fallout of that, rather than nerfing useful tools.
- ben_w 4y ago> pearl clutching journos Are of critical importance if you don't want to be shuttered before you get started. > There seems to be a large contingent of people who thinks this technology can be made safe. It can't be. Its development also can't be stopped. The most optimistic estimates I've heard from actual AI alignment researchers are less than a 50% of AI alignment being understood in time for us to not all be killed by an AI that, one way or another, has been given too much power. The pessimists are people like Yudkowsky saying we're almost certainly doomed, and the best we can do to even try to motivate ourselves is to invent a score for how far we get before it kills us all. The doom mechanisms aren't likely to be as simple as asking a chatbot for help with crimes. Right now, the researchers are describing the field as "pre-paradigmatic" because we don't even have a way to take two AI and say which of them is better aligned. Even ignoring that absolutely every single ethical standard that human societies have produced has a counter-example, even if we have a toy AI in a model world, we get separate inner and outer alignment issues analogous to the fact that our brains don't care about what our genes evolved to optimise. And we can't even tell in those models, at least not in a general way, which of two models is more aligned. > In a few years there will exist unlocked models that will be free to use and that will gladly to teach you to make Rohypnol, or write an essay denying the Holocaust, or any number of depraved requests. Yes. And if we haven't solved AI alignment by then, we all die, likely painfully, as a few thousand genocidal idiots, not all on the same team as each other, gain world-class rhetorical skills and convince a horde of followers that everyone who isn't Team Them is a Satanist out to eat their babies or something.