Loading...

Tag trends are in beta. Feedback? Thoughts? Email me at [email protected]

From NP-complete to O(N^2) to O(nlogn): Codegen strategies for case statements.

ArXiv's Updated Rate Limit Policy

Fixing GRPO's credit assignment problem without evaluating every step

Context Language Models

GPU-Initiated Communication: Dissecting Down to the Bone

Quantum information spreading via higher-order operator correlators

Harm Laundering in GPT Models: Gender Discrimination Transformed Rather Than

DeepSeek Elastic Compute (DSec)

LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes

The Implications of Linguistic Illegibility for LLM Security

AI Agents Push Humans Out of the Loop

Schedules Are Solvable Symbols: Tuning-Free Compilation of Tile Programs on Dataflow Architectures

LLMs as a Cognitive Virus

Thinking fast and slow in AI: The role of metacognition (2021)

Quantized Reasoning Models Think They Need to Think Longer, but They Do Not

Show HN: Training a model to identify AI web content from structure alone

"As a Language Model": Chat Template Switches LLM Self-Referential Voice

How good are frontier models at physics?

Dream-RSI: Recursive Self-Improvement through Evolving Worlds

Trusting-Trust Attack against an Entire Linux Distribution

ArXiv receives multiyear commitments to support it as an independent nonprofit

Compiler-style optimization for drawing via Skia

An empirical study of harness design for coding agents

Accurate Models of AMD Matrix Cores

Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data

Thinking with Looped Flows

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

Reflections on Trusting Trust, Revisited: Poisoning Self-Modifying AI Coding

The Malicious Use of Artificial Intelligence

Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation

More →