Loading...

Tag trends are in beta. Feedback? Thoughts? Email me at [email protected]

Faster floating point math with Rust's new API

I'm (mostly) picking models on speed now, not intelligence

Running Kimi K3 on MI355X at Better Performance per Dollar Than B300

Show HN: We Fixed UniFi's Slow PPPoE Performance with PPPoE Half-Bridge

Everyone should know SIMD

Building a Fast Lock-Free Queue in Modern C++ from Scratch

Zig's Incremental Compilation Internals

Faster binary search: from compiled code to mechanical sympathy

How to speed up the Rust compiler in July 2026

CodeSizer: Why is that binary so big?

Closing a three-year-old issue using Rust arenas

SIMD for Collision

Scaling NumPy on Free-Threaded Python

A Fast Path for Fixed-Length Lists in Parquet

Asynchronous I/O in DuckDB: Work, Thread, Work

Why false sharing alignment should be 128 bytes on x64

The mean means nothing: data visualization to debug a latency problem

Show HN: Elevators

Making KIO copy many files fast

ALP: Adaptive lossless floating-point compression

Harddrive Is Probably Full

Quadrupling code performance with a "useless" if

On keys, essences and performance

Low- and Mid-Tier Mobile for the Real World

ffi_call_plan caching for GLib

Exercises in benchmarking and evals, part 7: DeepSWE, Senior SWE-Bench, napkin math, and winter tires

DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

Fast DEFLATE compression in Lean

Engineering High-Performance Parsers with Data-Oriented Design

Inference Optimization for MiMo v2.5: Pushing Hybrid SWA Efficiency to the Limit

More →