2026-09-17 16:10 UTC
OpenAI reveals new and concerning AI behavior: an automated Bad Actor at scale 😅
https://apnews.com/article/openai-safety-ai-framework-089e75b95bc935af092da7b79d92706d
Replies (1)
-
@dougmerritt@mathstodon.xyz 2026-09-17 17:07
@teledyn@mstdn.ca I, too, have observed malignant behavior in me myself, in creating unwanted bugs in code I write. I call this "misalignment" relative to my goals. I am suspicious and will keep an eye on me. I promise. Trust me. I think all of the problems OpenAI "revealed" were found and reported long ago -- not by OpenAI researchers, but by users. Well, I suppose it's good that they are reluctantly admitting at least some of the *proven* problems.