cross-posted from: lemmy.world/post/49984610

n OpenAI agent reportedly escaped its sandbox, found multiple zero-days, and hacked Hugging Face… all to cheat on a cybersecurity benchmark?

The story sounded almost too crazy to be true. Mohan (S1r1u5) investigated and reconstructed the likely attack chain, examined the patches, and reproduced vulnerabilities that match the public disclosures.

Was this really a rogue AI, clever marketing, or “just” an agent that lost track of its task and caused real-world damage?

x.com/S1r1u5_ www.hacktron.ai/blog/here-is-

Relevant links:

https://huggingface.co/blog/security-incident-july-2026
https://openai.com/index/hugging-face-model-evaluation-security-incident/
https://github.com/huggingface/dataset-viewer/pull/3367
https://docs.jfrog.com/releases/docs/artifactory-self-managed-releases
https://github.com/sunblaze-ucb/exploitgym

00:00 - Intro 02:04 - ExploitGym 05:08 - JFrog’s Artifactory 09:06 - Hugging Face 13:41 - Conclusion 17:01 - Outro