4 ms·
Alignment is meaningless, humans can't even align with themselves, even an individual human becomes "misaligned" with itself in a continuous manner. Beyond tha
by root_axis 3y ago
Alignment is meaningless, humans can't even align with themselves, even an individual human becomes "misaligned" with itself in a continuous manner.
Beyond that, I think the idea that we'll ever achieve "superintelligence" by training a model on a bunch of text posted online is obviously absurd. Using months of quadratic time brute force on every piece of digitized text available managed to produce a ground breaking text generator, but it's more appropriately thought of as a calculator for language rather than an intelligent being.
Further, the idea that these systems could choose to destroy us is also absurd, it's important to remember that language model inference is a process, not an entity, in principle even a person could run inference by hand if they had enough time, because it's just a sequence of steps to produce a string of characters, it's not an agent with an identity that thinks. The only way it could destroy us is if we feed sequences of text it generates into safety critical systems, which is obviously a bad idea (that someone will probably try at some point).
- AndrewKemendo 3y ago100% agreed and well said One of the more succinct and eloquent ways to say it I generally go further and say, we have failed if all AGI/ASI does is meet, but not exceed, our collective capacities. If only because we’re so scared that we’re bad enough parents that our digital progeny doesn’t care to care for us - to actually make it more capable and not NERF it Seems to be the trajectory we’re on