Post #1844213
2026-02-23 21:06 UTC
This is Meta's head of AI "safety" doing more than making a “rookie mistake" as she put it.
The biggest failing here, imo, was not letting a probabilistic chatbot have full control over her email account. That's bad, but even worse is anthropomorphizing the LLM.
When it started to "misbehave" (remember, they don't have any real sense of true/false, and "rules" that you give them are just more text for the blender) she started pleading "don't do that" as though it was a person. It's a computer program. You can just turn it off.
This is part of the danger of using the intentionally misleading language that everyone has been pushed into using about these, such as (but definitely not limited to) "hallucination." Even people whose full time job is about this problem get tricked into forgetting what these are, and what they are not.
https://www.404media.co/meta-director-of-ai-safety-allows-ai-agent-to-accidentally-delete-her-inbox/
Replies (0)
No replies.