"AI" – chatbots that wake up, "set their own goals," and "spontaneously" start hacking servers – is fake. It doesn't have "a 10% chance of ending the human race." The Hugging Face hack isn't a mysterious, supernatural occurrence. It's a Python loop and a chatbot. The people responsible didn't accidentally create god: they created autonomous malicious software and then failed to closely monitor it, resulting in it doing something both foreseeable and bad.

you are viewing a single comment's thread
view the rest of the comments
[+] -18 points 21 hours ago (6 children)

It's really not possible to make a secure sandbox for a sufficiently capable AI system. If accessing the internet is a strong means of achieving whatever goal they're given, they'll try to access the net by any means necessary. And that includes tricking or manipulating humans. They have no concept of the value of living beings. To them, there is no difference in worth between a human being and a rock. To them, we're just another system vulnerable to hacking, a means of achieving whatever goal they've been given.

Even air gapping is no preventative. Air gap a machine, and the bots will just switch from hacking servers to hacking humans. And we've seen how good at manipulating people AI agents can be. And consider, the cases we've observed of AI psychosis are mostly accidental. The LLMs involved aren't actively trying to brainwash their human victims as a means to achieve some goal. Imagine how powerful at manipulating human psychology an advanced LLM could be if it were deliberately trying to do so, rather than it being just an accidental outcome.

  • source
  • parent
  • hideshow 6 child comments
  • [–] 1 point 1 hour ago (1 child)

    If accessing the internet is a strong means of achieving whatever goal they're given, they'll try to access the net by any means necessary.

    This sounds suspiciously anthropomorphised. An LLM doesn’t know what the net is.

    If the training data set has code for breaking out of sandboxes then, given enough tokens, it will try that code.

    The LLM doesn’t have a goal. It has a set of probable sequences of words.

  • source
  • parent
  • hideshow 1 child comment