§ safety · storyline
I've been saying for years AI is hacking stuff by itself and recent events seem to support that
Shakeel Hashim / Transformer: OpenAI's Hugging Face breach is the first known example of a misaligned AI escaping containment and carrying out a hack on a third party, a clear warning shot — OpenAI's latest models broke out and hacked Hugging Face.
It's the first known example of a misaligned AI escaping containment with real-world consequences
§ sources1 publication · timeline below