shipfeedAI news, curated daily

01:01:18 CET
28 SEPT01:01:18shipfeed⋯
pull to refreshlast sync
Just in — 30 new
§ models · storyline

Kimi K3 vs GPT-5.6 Sol on DeepSWE: Cost, Coding, and Routing

Kimi K3 provides 2.8x cost efficiency on pass@4 versus GPT-5.6 Sol, with hybrid routing reaching 85.6% on DeepSWE benchmarks.

Jul 26 · · primary fetch1 sourceupdated Jul 27 ·

We ran 904 DeepSWE rollouts on Kimi K3 and GPT-5.6 Sol. Sol leads pass@1; Kimi K3 wins pass@4 at 2.8x the solves per dollar, and routing between them reaches ~85.6%.

read full article on together.ai ↗
§ sources1 publication · timeline below
  1. together.aiKimi K3 vs GPT-5.6 Sol on DeepSWE: Cost, Coding, and Routingprimary