items50 latest
Mathematicians Build Long-Awaited Graph Sandwich
https://arxiv.org/abs/2510.20765
An Empirical Study of Harness Design for Coding Agents
Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data
Breaking the 1.58-bit Barrier for Ternary LLMs
How good are frontier models at physics?
Accurate Models of AMD Matrix Cores
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
Learning to solve hard problems in RL for LLMs by never giving up
https://arxiv.org/abs/2609.13443
The k-server conjecture is true
Backprop Alternative: Augmented Lagrangian Predictive Coding
https://arxiv.org/abs/2605.31022https://github.com/SakanaAI/pc-alm
The Malicious Use of Artificial Intelligence
Will There Be a 7G?
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
C*: Unifying Programming and Verification in C (2025)
Harnessing the Universal Geometry of Embeddings
LLMs as a Cognitive Virus
A mysterious kidney disease has arrived in Texas
The Emergent Symbolic Structure of Artificial Neural Networks
Longest Straight Line Paths on Water or Land on the Earth (2018)
Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment
Two German airport workers die of malaria after 'mosquito arrives on plane'
Small Models Have Arrived
mold: A Parallel Linker
Agentic Context Management: Memory and Cost as Architecture Problems
Black hole singularity is a surface not a point
France's tax agency got hacked (in French)
Public services are increasingly strained by LLM-written appeals for benefits
How I came to write that paper with Leslie Lamport
https://www.microsoft.com/en-us/research/publication/specifi...
Radiation damage to Hubble has been 4.3 years out of phase with the Solar cycle
DiffusionGemma Technical Report
Chain-of-Thought Reasoning in the Wild Is Not Always Faithful (2025)
Mathematics in the Age of AI
Code-native generation of highly programmable 3D assets (2026)
GPU Offload in Rust: Portable, Safe, and Fast
Launch HN: Speko (YC S26)
Hi HN! I'm Bek, founder of Speko, a platform that finds an optimal combination of speech-to-text, LLM, and text-to-speech models, given your constraints, among all our public benchmarked options, and tells you…