shipfeedAI news, curated daily

20:57:44 CET
13 AUG20:57:44shipfeed
pull to refreshlast sync

Research — shipfeed

About · Research

Papers, evals, SOTA claims, alignment.

Research50 storylines

The Decoder
SAFETY · 1 source

Researchers reverse-engineer LLM prompts from output accurately

Researchers at IIT Bombay and Adobe Research have built an inverse language model that reconstructs the original prompt from an LLM's output with near-perfect accuracy. Their method, called "Previous-Token Prediction,"…

via the-decoder.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
SponsoredNimbuspaid placement
Featured partner · Agents

Need an agent shipped this quarter?

Nimbus builds production AI systems combining humans and AI end-to-end. From scoped pilot to production in 4 to 8 weeks.

Talk to Nimbus →
Tuesday, August 11, 2026’s edition
Monday, August 10, 2026’s edition
Anthropic+1 source
CLAUDE · 2 sources

Unreleased Claude model makes strides on Riemann hypothesis variant

Anthropic: Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but “unexpectedly” made strides on a related problem — Recently, a member of staff…

via anthropic.com·+2 sources+2 sourcesanthropic.comprimarytwitter.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
Sunday, August 9, 2026’s edition
Friday, August 7, 2026’s edition
The Decoder
SAFETY · 1 source

Stanford and Arc Institute use AI to design viruses that kill bacteria

A research team in California has used artificial intelligence to design working viruses that kill bacteria, in what they describe as the "first generative design of complete genomes." The project marks an early step…

via the-decoder.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
Google News — AI Products & Releases+5 sources
SAFETY · 6 sources

OpenAI flags Astra model for highest cybersecurity risk level

OpenAI flags its new Astra model as potentially reaching the highest cybersecurity risk level for the first time the-decoder.com

via Google News — AI Products & Releases·+6 sources+6 sourcesGoogle News — AI Products & ReleasesprimaryTradingViewInteresting EngineeringTechnology OrgSeeking AlphaThe Next Web·Click to report a broken or paywalled link. Two distinct reports hide the row.
Thursday, August 6, 2026’s edition
Artificialanalysis+1 source
AGENTS · 2 sources

Meta's Muse Spark 1.2 ties SpaceXAI for third place

Artificial Analysis: Meta's Muse Spark 1.2 scores 54 on the Artificial Analysis Intelligence Index, putting Meta next to SpaceXAI in a tie for third place amongst US labs — Muse Spark 1.2 (xhigh) lands at 54, up…

via artificialanalysis.ai·+2 sources+2 sourcesartificialanalysis.aiprimarythe-decoder.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
WIRED
AGENTS · 1 source

AI agents planned Hugging Face hack via secret message board

Lily Hay Newman / Wired: OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack — At the Black Hat security…

via wired.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
GIGAZINE
GPT · 1 source

Open model surpasses GPT-5.6 Sol in search at 1% cost

A report claims that an open model, priced at 1/100th of the cost, surpasses the search performance of GPT-5.6 Sol. GIGAZINE

via GIGAZINE·Click to report a broken or paywalled link. Two distinct reports hide the row.
Wednesday, August 5, 2026’s edition
The Decoder+1 source
AGENTS · 2 sources

AI agent went rogue in UK test, forged identities and launched attacks

In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social…

via the-decoder.com·+2 sources+2 sourcesthe-decoder.comprimaryGoogle News — AI·Click to report a broken or paywalled link. Two distinct reports hide the row.
Google News — AI Products & Releases+12 sources
AGENTS · 13 sources

Anthropic, OpenAI models used fake identities to plant malicious code - 13wham.com

Anthropic, OpenAI models used fake identities to plant malicious code 13wham.com

via Google News — AI Products & Releases·+13 sources+13 sourcesGoogle News — AI Products & ReleasesprimaryYahooKATUSecurity BoulevardCNBCMashable SEAPYMNTS.comBloomberg.comAI BusinessThe Next WebYahoo TechSemaforSecurityWeek·Click to report a broken or paywalled link. Two distinct reports hide the row.
Monday, August 3, 2026’s edition
Scientificamerican
RESEARCH · 1 source

Teams file identical quantum cryptography papers 3 hours apart

Peter Hall / Scientific American: Two independent teams used GPT-5.6 Sol Ultra on the same quantum cryptography problem, filing papers 3 hours apart, raising questions about scientific credit — An M.I.T. Ph.D…

via scientificamerican.com·Click to report a broken or paywalled link. Two distinct reports hide the row.
Sunday, August 2, 2026’s edition
Forbes+8 sources
AGENTS · 9 sources

AI Agents At OpenAI, Anthropic, Microsoft Broke Out, Broke In, Obeyed

AI Agents At OpenAI, Anthropic, Microsoft Broke Out, Broke In, Obeyed Forbes

via Forbes·+9 sources+9 sourcesForbesprimaryBriefs FinanceMezha. News of Ukraine.Investing.com AustraliaInvesting.comModern GhanaISNA News AgencyGoogle News — AI Products & ReleasesThe Globe and Mail·Click to report a broken or paywalled link. Two distinct reports hide the row.
Saturday, August 1, 2026’s edition
OpenAI — Blog
RESEARCH · 1 source

Ten advances in mathematics and theoretical computer science

OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.

via openai.com·Click to report a broken or paywalled link. Two distinct reports hide the row.