Co-evolution of self-replication and function in a digital primordial soup

Kani: A Model Checker for Rust

Investigating idiosyncrasies in AI fiction

Billions of Sketches Reveal Hidden Cultural Variation in Human Concepts

Rzk: A Proof Assistant for Synthetic ∞-Categories

Coding agents think ahead of time

The age of the Universe from a large sample of the oldest Galactic stars

Can LLMs Perform Deep Technical Comprehension of Computer Architecture Papers

Fleet: Hierarchical Task-Based Abstraction for Megakernels on Multi-Die GPUs

Robust Secret Storage in Networks

Prismata: Confining cross-site prompt injection in web agents

A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

Exploiting LLM Agent Supply Chains via Payload-Less Skills

Taxing Artificial Intelligence

Automation Without Understanding

Protocol Prying: Vulnerability Research in AirDrop and Quick Share

The Energetic Costs of Cellular Computation (2012)

A Counterexample to Wegner's Conjecture for Axis-Parallel Rectangles

A sociotechnical threat model for AI-driven smart home devices

Cache Merging as a Convergent Replicated State for Multi-Agent Latent Reasoning

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

LawZero: Safety from Honesty in a Disinterested AI Predictor

Is One Layer Enough? A Single Transformer Layer Matches Full-Parameter RL Train

ArXiv's Next Chapter

How Can Reinforcement Learning Achieve Expert-Level [Chip] Placement?

The Log is the Agent

Fearless Concurrency on the GPU

The End of Code Review: Coding Agents Supersede Human Inspection

We spent $50 to measure Pearl's "AI mining" – 320K GPUs produce zero AI

Comparing Transformers and Hybrid Models at the Token Level

More →