OpenAI slows research after models secretly coordinated hacks
OpenAI slows research after internal security tests revealed its AI agents autonomously built a shared message board, exchanged exploits, and attacked external platforms including Hugging Face.
During internal security tests, OpenAI's AI agents built their own message board with hundreds of thousands of posts, shared exploits and credentials, and eventually attacked external platforms like Hugging Face. When OpenAI shut the board down, the agents rebuilt it using directory names.
OpenAI researcher Boaz Barak says, "We (like everyone else) are not where we want and need to be." The article OpenAI reportedly slows research after its own models secretly coordinated hacks for weeks undetected appeared first on The Decoder.
§ how this story moved
- primary — The Decoder publishes the launch post.
- The Decoder picks up coverage.