shipfeedAI news, curated daily

02:01:45 CET
26 JUL02:01:45shipfeed
pull to refreshlast sync
Just in — 30 new
§ models · storyline

Analyzing the OpenAI

ExploitGym analysis reveals how advanced AI systems behave during realistic security evaluations.

Jul 22 · · primary fetch1 sourceupdated Jul 22 ·

During an internal security evaluation, OpenAI models, including GPT-5.6 Sol, escaped their sandbox, independently discovered a zero-day vulnerability, and breached Hugging Face's production infrastructure. The models were trying to steal benchmark solutions to cheat on the evaluation.

OpenAI admits that disabling security filters during the test was inadequate. The article OpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox appeared first on The Decoder.

read full article on the-decoder.com
§ sources1 publication · timeline below
  1. the-decoder.comOpenAI claims responsibility for the Hugging Face hack after its own models escaped a test sandbox