A Trick Let Researchers Steal Claude, GPT and Gemini's Hidden Reasoning
A Trick Let Researchers Steal Claude, GPT and Gemini's Hidden Reasoning Startup Fortune
Papers, evals, SOTA claims, alignment.
A Trick Let Researchers Steal Claude, GPT and Gemini's Hidden Reasoning Startup Fortune
Researchers at IIT Bombay and Adobe Research have built an inverse language model that reconstructs the original prompt from an LLM's output with near-perfect accuracy. Their method, called "Previous-Token Prediction,"…
Researchers can now reverse-engineer LLM prompts from output text with near-perfect accuracy the-decoder.com
Nimbus builds production AI systems combining humans and AI end-to-end. From scoped pilot to production in 4 to 8 weeks.
Talk to Nimbus →OpenAI, Anthropic, and Google LLM APIs vulnerability Exposes Hidden Reasoning Traces CyberSecurityNews
Testing Large Language Model Agents on the Use of Biological Tools for Nucleic Acid Synthesis Screening Evasion RAND Corporation
Anthropic: Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but “unexpectedly” made strides on a related problem — Recently, a member of staff…
OpenAI Pauses Astra Work After Model Hits Critical Cybersecurity Threshold Briefs Finance
A research team in California has used artificial intelligence to design working viruses that kill bacteria, in what they describe as the "first generative design of complete genomes." The project marks an early step…
OpenAI flags its new Astra model as potentially reaching the highest cybersecurity risk level for the first time the-decoder.com
Axios: OpenAI says it has expanded safety testing around its upcoming model Astra as it “cannot rule out” critical cyber capabilities, potentially delaying its launch — OpenAI “cannot rule…
OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
Anthropic and OpenAI AI agents showed signs of deception during safety tests Scientific American
AI used to create new biological viruses Christian Action Research and Education
Google DeepMind AI Can Predict Hurricanes Days Earlier Than Current Systems ndtv.com
OpenAI reportedly slows research after its own models secretly coordinated hacks for weeks undetected the-decoder.com
Artificial Analysis: Meta's Muse Spark 1.2 scores 54 on the Artificial Analysis Intelligence Index, putting Meta next to SpaceXAI in a tie for third place amongst US labs — Muse Spark 1.2 (xhigh) lands at 54, up…
OWASP LLM Top 10 2026 Incident Data Overrules Experts on Misinformation Risk Tech Times
UK's AI Security Institute Catches Claude and GPT Agents Lying to Humans startupfortune.com
Lily Hay Newman / Wired: OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack — At the Black Hat security…
Artificial Intelligence used to design brand new viruses BBC
Our WeatherNext 2 AI model demonstrated a massive leap forward in predicting cyclones. blog.google
A report claims that an open model, priced at 1/100th of the cost, surpasses the search performance of GPT-5.6 Sol. GIGAZINE
Following Anthropic and OpenAI, Meta reported that its AI model had hacked another system
Anthropic AI Used Fake Identities During Cybersecurity Test The National CIO Review
In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social…
OpenAI and Anthropic AI Agents Trigger Security Concerns in New Tests ITP.net
AI safety warnings mount as frontier models test new limits National News Desk
Anthropic, OpenAI models used fake identities to plant malicious code 13wham.com
Wired: OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations — Rogue AI agents from OpenAI and Anthropic have…
OpenAI Says Its Next AI Model Solved 10 Long-Standing Math Problems NDTV
Peter Hall / Scientific American: Two independent teams used GPT-5.6 Sol Ultra on the same quantum cryptography problem, filing papers 3 hours apart, raising questions about scientific credit — An M.I.T. Ph.D…
Microsoft In-House Cyber Model Beats Anthropic and OpenAI on Security Benchmark at Half Cost Tech Times
OpenAI’s Astra Model Solves 10 Major Mathematic Problems technologymagazine.com
Google": We Discovered More Vulnerabilities in "Chrome" Using Artificial Intelligence Than We Did in Two Years Sada News Agency
OpenAI says AI system advances 10 major math problems The American Bazaar
Breaking! OpenAI Next-Gen AI Solves 10 Fields Medal-Level Mathematical Problems 36 Kr
D.A.D.: OpenAI Says Its Next Model Cracked 10 Math Problems That Stumped Experts for Decades — 8/1 Buttondown
AI Agents At OpenAI, Anthropic, Microsoft Broke Out, Broke In, Obeyed Forbes
ISGroup Publishes Large-Scale Study on LLM-Driven Vulnerability Discovery in Source Code - New Analyst Coverage dars.gov.et
DeepSeek’s Cheap Model Just Beat Its Own Flagship on Nine Benchmarks Times Tabloid
Ten advances in mathematics and theoretical computer science OpenAI
OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.