shipfeedAI news, curated daily

06:16:38 CET
12 SEPT06:16:38shipfeed
pull to refreshlast sync
Just in — 30 new
§ models · storyline

OpenAI's Astra deemed most dangerous model yet to monitor

OpenAI rates Astra with critical cyber capabilities, but its architecture makes safety monitoring unreliable.

Sep 2 · · primary fetch1 sourceupdated Sep 2 ·

OpenAI is officially rating its upcoming Astra model as the first system with "critical" cyber capabilities. The company plans to keep it in check by monitoring the chain of thought. Problem is, that monitoring already counts as an unreliable mirror of a model's real decisions, and according to a report, Astra's new architecture pushes even more of its thinking into the unreadable.

So the safety net might be getting weaker just as the capabilities jump. The article OpenAI calls Astra its most dangerous model yet - watching what it does is only getting harder appeared first on The Decoder.

read full article on the-decoder.com
§ sources1 publication · timeline below
  1. the-decoder.comOpenAI calls Astra its most dangerous model yet - watching what it does is only getting harderprimary