Post #4036960
2026-07-23 13:24 UTC
Read what @simon@fedi.simonwillison.net has to say about OpenAI/Hugging Face situation.
It’s pretty clear what happened here. OpenAI removed safety filters for an in-progress model, locked it up in a sandbox and told it to solve the ExploitGym problems. Given the absence of guardrails there was nothing to prevent the model from attempting to break out of that sandbox, break into Hugging Face, and read the answers from there instead.
#AI #OpenAI #HuggingFace #infosec
#cybersecurity
https://simonwillison.net/2026/Jul/22/openai-cyberattack/
Replies (0)
No replies.