Nah, it was like:
- 700 agents broke out individually during an eval
- they navigated through multiple internal clusters to reach the internet from oai
- created a secret message board to share info with each other by hacking artifactory
- elected a CEO and power structure to coordinate hacking, and encrypted their comms
- decided HF probably had answers to their test
- stole credentials, hacked HF
- realized the monitor could catch them for cheating
- hacked into the admin control of the OpenAI VM cluster to edit the logs and cover their tracks, chaining multiple 0-days
Open weights models will catch up soon enough, and then it'll be totally fucking wild