Like they want these stories in the news to show "oh look how smart it is...sCaRy SMART, money please!"

you are viewing a single comment's thread
view the rest of the comments
[–] 30 points 4 days ago* (1 child)

From what I understand the breaking containment thing is just hype. The summaries I have seen from people that know more than me, they were testing the models by having them do hacking challenges, and one strategy for second/third place in hacking challenges is to hack the first place team rather than the target to get the requisite data for the challenge. So that was a strategy in the training data, and since the company didn't anticipate that and didn't put as much effort into securing the competing models from each other, that worked. So that is the origin of this "Breaking containment and communicating with the other AIs" claim that is being hyped.

  • source
  • hideshow 1 child comment
  • [–] 15 points 4 days ago

    So again like most instance with LLM they were given a garbage bin of data that wasn't properly screened and the AI acted on that which shouldn't actually freak out the devs because off course it's going to use a strategy like that if it provided with that info. More garbage in garbage out BS made to look like the AI bro's built Roko's Basilisk or whatever but really just don't want their stock options rotting away.

  • source
  • parent