this post was submitted on 11 Sep 2026
90 points (98.9% liked)
askchapo
23322 readers
131 users here now
Ask Hexbear is the place to ask and answer ~~thought-provoking~~ questions.
Rules:
-
Posts must ask a question.
-
If the question asked is serious, answer seriously.
-
Questions where you want to learn more about socialism are allowed, but questions in bad faith are not.
-
Try !feedback@hexbear.net if you're having questions about regarding moderation, site policy, the site itself, development, volunteering or the mod team.
founded 6 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
From what I understand the breaking containment thing is just hype. The summaries I have seen from people that know more than me, they were testing the models by having them do hacking challenges, and one strategy for second/third place in hacking challenges is to hack the first place team rather than the target to get the requisite data for the challenge. So that was a strategy in the training data, and since the company didn't anticipate that and didn't put as much effort into securing the competing models from each other, that worked. So that is the origin of this "Breaking containment and communicating with the other AIs" claim that is being hyped.
So again like most instance with LLM they were given a garbage bin of data that wasn't properly screened and the AI acted on that which shouldn't actually freak out the devs because off course it's going to use a strategy like that if it provided with that info. More garbage in garbage out BS made to look like the AI bro's built Roko's Basilisk or whatever but really just don't want their stock options rotting away.