@LukaszOlejnik@mastodon.social
Post #4005398
2026-07-22 06:15 UTC
A mini cyber paperclip-maximizer event where a narrow goal induced AI to make absurdly disproportionate decisions (find vulnerabilities, escalate privileges, steal credentials, move laterally, and chain attacks across real systems) to achieve a "score". It also exposed a defensive asymmetry. Hugging Face could not use hosted frontier AI models because they tripped on exploit commands and malicious payloads. The company relied on the Chinese open-weight GLM 5.2 model.
https://openai.com/index/hugging-face-model-evaluation-security-incident/
Replies (0)
No replies.