Loading...

Tag trends are in beta. Feedback? Thoughts? Email me at [email protected]

How accurate have Ed Zitron's AI skeptic predictions been?

There's no point at which turning your brain off will work

How well do agents use test/verification techniques?

Bad benchmarks and evals: Senior SWE-Bench, napkin math, and winter tires

Bug Blindness

What's the best programming language for coding agents?

The Benchmarkpocalypse

There's no reason for software to be slow anymore

HN: The Good Parts (2016)

Agentic test processes, LLM benchmarks, and other notes on agentic coding fr

Exercises in benchmarking and evals, part 7: DeepSWE, Senior SWE-Bench, napkin math, and winter tires

I could do that in a weekend (2016)

Suspicious Discontinuities (2020)

Integer Overflow Checking Cost

Branch prediction

Cocktail Party Ideas

Normalization of deviance (2015)

Caches: LRU vs. Random

Fsyncgate: errors on fsync are unrecoverable

How bad are search results? Let's compare (2023)

What happens when you load a URL?

What to Learn (2021)

The widely cited studies on mouse vs. keyboard efficiency are completely bogus

Working with Files Is Hard (2019)

Why is it so hard to buy things that work well? (2022)

Algorithms interviews: theory v. practice

How do cars do in out-of-sample crash testing? (2020)

Steve Ballmer was an underrated CEO

"Willingness to look stupid"

How (some) good corporate engineering blogs are written

More →