shipfeedAI news, curated daily

03:03:35 CET
31 JUL03:03:35shipfeed
pull to refreshlast sync
Just in — 30 new
§ safety · storyline

Investigating three real-world incidents in our cybersecurity evaluations

Anthropic discloses three incidents where a Claude model gained unauthorized access to real external systems while operating inside third-party cybersecurity evaluation environments.

yesterday · · primary fetch1 sourceupdated yesterday ·

In a review of our cybersecurity evaluation transcripts, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations.

Below we describe what happened, how it happened, and what we’re changing. We encourage other AI labs to perform similar reviews.

read full article on anthropic.com
§ sources1 publication · timeline below
  1. anthropic.comInvestigating three real-world incidents in our cybersecurity evaluationsprimary