

Yeah, I think that’s where the real crime is: the fact that OpenAI was asleep at the wheel when it came to security and monitoring. This should not have happened at all, but they didn’t even realize it even happened in the first place until many days after the fact. (I’ve also heard that most, if not all, of these incidents involve the company in question outsourcing their sandboxing (??) to some “frontier AI security” startup called Irregular. link)



Interesting viewpoint. Honestly, this serves as a great example of how the idea of evolution is actually more subtle than many people think. With biological evolution, most people are taught not to assign any intention to the process, but here we have people reifying the AI agents with desires to break out of the sandbox and cheat on the assignment.
Another part of it is people severely underestimate just how many resources the AI companies have to spend on stunts like this. How many millions of dollars they spend running hundreds of millions of dollars of hardware to perform this little experiment for several weeks? Perhaps this makes evolution a better viewpoint than bad actors with intentions.