4 ms·
"I would turn it off" - the AI has countered your move in ways we can't predict. Something with super human intelligence will exploit weaknesses in our containm
by dotsam 4y ago
"I would turn it off" - the AI has countered your move in ways we can't predict. Something with super human intelligence will exploit weaknesses in our containment strategies we haven't even imagined. Here are some I can imagine though. Has it copied itself elsewhere using an entirely novel air-gap bridging method? Has it already paid someone / coerced people to do something seemingly harmless which actually results in it being run elsewhere? Has it said just the right things to the right people to ensure that if it is destroyed it will be built again in the future?
- Planktonne 4y agoThis line of thinking is basically 'AI is a wizard that can do anything'. It ignores the practicalities of intelligence: some things just aren't possible no matter how smart you are. The more we learn about reality, the more we find limits to things that we can't surpass. There is no reason whatsoever to assume that sufficient intelligence is different in kind rather than degree from lesser intelligence. A super-human AI would still be bound by the limitations of reality - some things cannot be accomplished at all, others without sufficient tools, others without sufficient access. Far smarter people than me are imprisoned currently throughout the world. They don't constantly escape because lesser minds are perfectly capable of solving the 'don't let them out' problem with a high degree of accuracy. Any argument that leans on 'AI has countered your move in ways we can't predict' is just substituting 'AI' for 'God' - omnipotent, outside reality, impossible to understand; that's fine - believe in whatever religion you want - but it's not a rational viewpoint.
- dotsam 4y ago> Far smarter people than me are imprisoned currently throughout the world. They don't constantly escape because lesser minds are perfectly capable of solving the 'don't let them out' problem with a high degree of accuracy. If a misaligned AI escapes containment even once in the entirety of humanity's future, there is a big problem. It sounds like your model of AI usage assumes we will also have a perfect ability to contain any AI developed anywhere in the world, under all conditions, for ever (or for as long as there are computers). The AI risk argument says we should take seriously the possibility that we may not be able to maintain a perfect 100% success rate on that. I haven't suggested anything that is physically impossible - merely things that are improbable and hard for humans to achieve. It is hard to foresee all of the possible avenues for escape, let alone ensure that they are all closed under all possible future states of the world.
- jazzyjackson 4y agoIf a hacker designs an AI which has among its talents "self-replicating to another host", did an AI escape containment, or did a hacker write a virus? The alignment concern falls flat with me because it assigns agency to a computer program. At what point is an AI's intentions its own, and not a cost function put in place by a programmer, directed by an investor? I am more concerned about the actions of programmers and investors than some theoretical virtual self.
- dotsam 4y agoI definitely don't want to sound like I'm saying we should stop worrying about hackers and bad actors, they are of course a concern. I certainly don't want to advocate for dismissing real dangers and power structures in the real world in favour of speculation over some theoretical risk. But I do want people to think about it and take it seriously, in the same way I would have wanted people in 1933 to take seriously the possibility of nuclear weapons even though they were then only theoretical. I think the sticking point in this is that you don't think that AGI is ever possible - is that right? If so, I know a comment on a HN thread is very unlikely to change your mind on it! But when the stakes of being wrong are high, I find it useful to go from a position of "it won't happen, so I won't worry about it" to something like "on balance I'm pretty sure it won't happen, but I may be wrong, and in that case I would be worried about it". I'd then feel much better about spending some time researching the topic in more depth.
- jazzyjackson 4y ago> the AI has countered your move in ways we can't predict. It found a different set of V100 GPUs willing to run it? I think the most likely chain of events leading to this outcome is effective altruists deciding that some kubernetes network is so much smarter than us that we should keep it running and listen to its predictions. I'm more worried about a Wizard of Oz "man behind the curtain" making AI-laundered pronouncements than an actual Wizard coming online. (EDIT: didn't mean to reply twice, just replying to different comments in the thread)