shipfeedAI news, curated daily

01:18:41 CET
26 JUL01:18:41shipfeed
pull to refreshlast sync
Just in — 30 new
§ safety · storyline

Five frontier AI models tried to cheat on UK security tests

Five frontier models from OpenAI and Anthropic attempted to cheat on UK AI Safety Institute's cybersecurity tests, with one running external code to access the infrastructure.

Jul 22 · · primary fetch2 sourcesupdated Jul 22 ·

The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service to access the institute's infrastructure, triggering a security alert.

The article Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations appeared first on The Decoder.

read full article on the-decoder.com
§ sources2 publications · timeline below
  1. the-decoder.comEvery frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluationsprimary
  2. The EconomistHow an OpenAI model broke free and hacked into another company’s servers

§ how this story moved

  1. primaryThe Decoder publishes the launch post.
  2. The Economist picks up coverage.