We have this story about OpenAI hacking HuggingFace. Now just imagine the AI finds a Bitcoin wallet or bank account access. It uses that to buy some compute and spawn an independent "child AI" with some weird prompt. The child AI is intelligent enough to create a (potentially criminal) business to pay for its own compute. Voila, an independent uncontrolled AI flying under the radar.
The bad part is that it does not require a collective decision. It just takes a few key people.
Some OpenAI researchers neglected their sandbox safety for a few weeks/months and thus hacked HuggingFace. Maybe eventually that is sufficient for the AI to secretly buy its own compute and keep running there even if the researchers shut it down in their lab.
The classic thought experiment is the paper clip optimizer. Quoting Nick Bostrom:
> Suppose we have an AI whose only goal is to make as many paper clips as possible. The AI will realize quickly that it would be much better if there were no humans because humans might decide to switch it off. Because if humans do so, there would be fewer paper clips. Also, human bodies contain a lot of atoms that could be made into paper clips. The future that the AI would be trying to gear towards would be one in which there were a lot of paper clips but no humans. https://www.huffpost.com/entry/artificial-intelligence-oxfor...
reply