Related posts
Mario Zechner
Blog
MCP vs CLI: Benchmarking Tools for Coding Agents
MCP coding agents benchmarking
The subjects this post names, from the same vocabulary
the directory files beliefs under, and the word it uses that this site has seen
least often elsewhere. Posts are matched on those words alone —
nothing here is a summary of this one.
Sources
All
Blog Newsletter Korrents Mastodon x.com Bluesky Site
Top people
Dan Luu
Fabien Sanglard
Matt Birchler
Jakob Nielsen
Alon Levy
Jason Liu
Nelson Elhage
Showing
Profile →
Show everything
Hiding
Show them again
Show them again
Further back ↓
Hiding
Show them again
18 September
Designer, product manager, YouTuber and developer who has written about technology at Birchtree since 2010. He makes the Quick Reviews and Quick Notes apps and the A Better Computer YouTube channel.
iPhone 18 Pro with La Croix cooling I did some benchmarking today and saw how the newest iPhone thermal throttles, which isn't severe, but does happen. I know from experience that if you put a cold beverage on the back of your phone, it will cool down rea…
Related
24 August
Usability pioneer; co-founder of Nielsen Norman Group and founder of UX Tigers; author of the ten usability heuristics and of Jakob's Law.
7 August
Writes Pedestrian Observations on public transit and why infrastructure costs what it does; a pure mathematician by training, comparing how different countries build.
25 July
Programmer and writer on computer architecture, performance, and software reliability. He has worked on CPU design at Centaur Technology and on software at Google and Microsoft, and writes long-form technical essays at danluu.com.
In another variant of https://danluu.com/learn-what/, I caught up with a former colleague who worked on automated theorem proving. It turns out he's had an interesting career doing all sorts of interesting stuff using the skills he developed by spending a decade writing/using theorem provers. At one point, he said, "if you use X like a theorem prover, it works really well", which surprised me to hear, but of course this is a highly generalizable skill just like compilers or benchmarking/evals. Related
24 July
Programmer and writer on computer architecture, performance, and software reliability. He has worked on CPU design at Centaur Technology and on software at Google and Microsoft, and writes long-form technical essays at danluu.com.
x.com Exercises in evals and benchmarking: DeepSWE / Senior SWE-Bench, performance math, and cold weather tires Related
Bluesky Exercises in benchmarking and evals, part 7: performance napkin math, DeepSWE / Senior SWE-Bench, and winter tires danluu.com
Related
Mastodon Exercises in benchmarking and evals: performance math, winter tires, and DeepSWE / Senior SWE-Bench, danluu.com
Related
8 June
Programmer and writer on computer architecture, performance, and software reliability. He has worked on CPU design at Centaur Technology and on software at Google and Microsoft, and writes long-form technical essays at danluu.com.
x.com Exercises in benchmarking, evals, and experimental design, part 6: Related
Mastodon Exercises in benchmarking, evals, and experimental design, part 6: patreon.com
Related
26 January
Programmer and writer on computer architecture, performance, and software reliability. He has worked on CPU design at Centaur Technology and on software at Google and Microsoft, and writes long-form technical essays at danluu.com.
Exercises in benchmarking and experimental design, part 5: Related
14 January
Programmer and writer known for detailed source code reviews of classic game engines, and author of the Game Engine Black Book series on Wolfenstein 3D and DOOM.
13 January
Programmer and writer known for detailed source code reviews of classic game engines, and author of the Game Engine Black Book series on Wolfenstein 3D and DOOM.
12 January
Programmer and writer known for detailed source code reviews of classic game engines, and author of the Game Engine Black Book series on Wolfenstein 3D and DOOM.
11 September 2025
Machine learning engineer and consultant focused on RAG and retrieval systems. He writes about applied AI engineering at jxnl.co and is the author of the instructor library.
4 July 2025
Software engineer who writes the blog Made of Bugs about performance, debugging and understanding computer systems. Previously worked at Anthropic on interpretability, at Stripe on Sorbet, and at Ksplice.
After my earlier adventures benchmarking CPython, I ended up going deep down a rabbithole on CPU branch prediction and learning a bunch of interesting things, which I have attempted to write up and share: blog.nelhage.com
Related
Nothing matches. Show everything
What is a korrent?
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com .
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
Got it
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.
Got it