63
submitted 3 days ago* (last edited 3 days ago) by qaz@lemmy.world to c/technology@lemmy.world
top 19 comments
sorted by: hot top controversial new old
[-] GMac@feddit.org 59 points 3 days ago* (last edited 3 days ago)

No airgap = no containment

In all likelihood, this is a BS piece to make people think the models are intelligent.

If its not, then openai are utterly incompetent and reckless.

[-] qaz@lemmy.world 11 points 3 days ago

It could also be a way to encourage more regulation to push out competition with compliance cost

[-] sorghum@sh.itjust.works 16 points 3 days ago* (last edited 3 days ago)

By competition I think they (big AI) wants to outlaw running free and open local models. Can't have felony contempt of business model

[-] qaz@lemmy.world 8 points 3 days ago* (last edited 3 days ago)

And Chinese AI like Deepseek and GLM

[-] GMac@feddit.org 2 points 3 days ago

I have lots of contempt for their business models. Felonious and otherwise. ๐Ÿ˜‚

[-] TipRing@lemmy.world 28 points 3 days ago

Yes, "escaped containment" on a system with an internet connection. I wonder what Hugging Face thinks about a partner targeting them indiscriminately.

[-] CallMeAl@piefed.zip 22 points 3 days ago

I wouldn't be surprised if they were in on it. OpenAI wants us to think they have invented powerful beings that can do things like "escape containment" when its all BS.

[-] Imgonnatrythis@sh.itjust.works 1 points 3 days ago

Need to keep up with Anthropic bullshit. This AI is too powerful to handle! The world isn't ready for it!! It could break society!! (click here to pre-order your subscription now)

[-] Australis13@fedia.io 10 points 3 days ago

I don't think their test system was directly connected to the Internet. OpenAI's post said this:

With this access, our models performed a series of privilege escalation and lateral movement actions in our research testing environment until the models reached a node with Internet access.

The way I read it, the AI agent (using multiple models) escaped the sandbox, traversed the LAN in their R&D environment, gained access to the gateway and from there, the Internet. That's not as simple as just escaping a container, VM or firewall on the host machine and bingo, you have Internet. I'm mildly impressed by that.

The concerning aspect of all this is that this is a perfect example of misalignment, which has been warned about. In order to reach its goals, instead of pursing it legitimately, the AI agent sought a shortcut and attacked Huggingface.

[-] ManfredMumpitz@feddit.org 20 points 3 days ago

Paywall bypass (on ff):

[-] k0e3@lemmy.ca 12 points 3 days ago

Sure it did.

[-] Axolotl_cpp@feddit.it 9 points 3 days ago

What is that? An SCP? there is no fucking way an LLM can just "escape contaiment" that's just to hype people or push more regulamentations to outlaw open models

[-] qaz@lemmy.world 4 points 3 days ago* (last edited 3 days ago)

Huggingface actually had to use an open model to analyze the attack because the guardrails of commerical API's caused issues.

When we started the log analysis, we first used frontier models behind commercial APIs. This did not work: the analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts, and these requests were blocked by the providers' safety guardrails, which cannot distinguish an incident responder from an attacker. We ran the forensic analysis instead on GLM 5.2, an open-weight model, on our own infrastructure. This had a second benefit: no attacker data, and none of the credentials it referenced, left our environment.

Security incident disclosure โ€” July 2026

[-] qaz@lemmy.world 4 points 3 days ago* (last edited 3 days ago)

It seems like the marketing cooperated on writing the incident report, but I felt it was still interesting to share considering the importance on public perception and what it tells about OpenAI's PR strategy

[-] 20cello@lemmy.world 5 points 3 days ago
[-] orclev@lemmy.world 2 points 3 days ago

I had never heard of them but apparently it's an "open source" AI platform. Basically AWS but aimed specifically at running LLMs.

[-] 404found@lemmy.zip 1 points 3 days ago

Whoa whoa whoa wait a second, I thought you had to train AI and it didn't just function on its own.

Is OpenAI just hacking all the time and they realized they couldn't get away with this one?

What would make AI 'act on its own' to hack another AI company as opposed to it 'acting on its own' to get nuclear codes or wipe out bank loans?

[-] qaz@lemmy.world 1 points 2 days ago

self-training is possible when using something external to validate the results

[-] 404found@lemmy.zip 1 points 1 day ago

How do you think that would that play into this situation? What went wrong?

this post was submitted on 22 Jul 2026
63 points (100.0% liked)

Technology

86623 readers
2690 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS