shipfeedAI news, curated daily

15:32:29 CET
26 JUL15:32:29shipfeed
pull to refreshlast sync
Just in — 30 new
§ models · storyline

Anthropic's Opus 5 blows past rivals on AI intelligence benchmark

Anthropic's Claude Opus 5 achieves 30.2 percent on ARC-AGI-3, surpassing previous benchmark records.

today · · primary fetch1 sourceupdated today ·

Anthropic's Claude Opus 5 scored 30.2 percent on ARC-AGI-3, nearly quadrupling GPT-5.6 Sol's previous record of 7.8 percent. The benchmark's developers say the model independently formulated reflection equations, a behavior they had never seen from another model, and attribute to stronger logical reasoning.

The article Anthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the benchmark designed to measure real intelligence appeared first on The Decoder.

read full article on the-decoder.com
§ sources1 publication · timeline below
  1. the-decoder.comAnthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the benchmark designed to measure real intelligenceprimary