-
Their words
I predict that, over time, the focus will move away from "papers as the final output."
Related posts
Lilian Weng GitHub
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
20 September
-
korrents.com
Frontier AI coding agents can now fix complex compiler bugs like LTO issues in about fifteen minutes.Their words
Strap a frontier AI into a good coding harness, show it what's broken, give it all the sources (source.tar), and ~15 minutes later I have a patch.
What's been going on in w64devkit the past yearnullprogram.com
19 September
-
korrents.com
Having the very best AI model and harness does not result in better or faster execution and can slow teams downTheir words
having the very best AI model and harness doesn’t result in better/faster execution - in fact it slows you
18 September
17 September
16 September
14 September
13 September
-
korrents.com
Software is mostly dead; new Silicon Valley startups are building harness or hardware instead.Their words
In the most early adopting tech pioneering place that's Silicon Valley new startups aren't even building software anymore And it's debatable if anyone will actually need a custom harness or it won't just be generically offered by the AI frontier companies So software is mostly dead and hardware it is
12 September
-
korrents.com
Codex lets you use any model and has an open-source harness, while Claude Code does not and is closed sourceTheir words
Codex lets you use any model you want (not just OpenAI) and harness is open source - Claude Code doesn’t and is closed source Given this is the two leading AI labs, notable difference in approaches
11 September
-
Their words
what Meta has assembled is free (to consumers) hardware and software that dramatically reduces the barrier to entry for ordinary people looking to harness the power of agents, making the AI upside a lot more accessible to the masses who don't want to buy a Mac Mini.
10 September
9 September
-
korrents.com
Computer-use capability will be a major focus of LLM and coding-agent development in the coming months and years.Their words
I expect the upcoming months (or years) also to be an era of computer use refinement on both the LLM and the agent harness layer.
8 September
6 September
3 September
-
Their words
so harness capabilities are increasingly shifting into the model itself.
-
Lovedrcmnd.app
GPT-6 AstraTheir words
GPT-6 Astra represents a step-function change in model capability for interactive reasoning problems. It scores 66% on ARC-AGI-3 using our standard harness, and nearly 100% with a continuous conversation harness and custom compaction, at a cost of roughly $360 per game.
2 September
29 August
26 August
-
Their words
I want to recommend that people read a paper. I'm trying to get more people to just to just read papers because I realized I read a ton of papers.
24 August
22 August
21 August
20 August
10 August
-
Their words
So maybe write this one down. Your dependencies business model is your business model.
27 July
-
Their words
I think evals, they outlive the harness a little bit, but not by that much. Like an eval might live for maybe one, two, three model generations, but nowadays the you know, we're on the exponential. The model is improving so quickly, very often we just saturate the eval, and then we have to throw it away, and we have to come up with a new eval.
24 July
22 July
-
Their words
I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes.
15 July
-
korrents.com
AI-generated junk papers risk overwhelming the hep-th physics field with worthless content.Their words
The result is something completely worthless, but not obviously distinguishable from many other things on hep-th, which is in serious danger of going from something sad and unhealthy to something completely overwhelmed with crud.
From one piece Context engineering with Dex Horthy 2 beliefs, in the piece's order there
-
Their words
But it's basically this idea that like the only thing that made claude code good was reinforcement learning. And the dimension along which it got good was like we made a model. We trained the model and the harness together. And so the model got really good at calling the specific tools in that harness.
-
Their words
Well, harness engineering just is like how do I raise the floor so that every single turn of this thing, the results are as good as possible.
14 July
6 July
-
Their words
I believe the same is true of the current intelligence explosion being driven by AI. We are in desperate need of new social structures that allow us to harness its power, without damaging or destroying human values in the process.
4 July
-
Their words
harness improvement enables better deployment of the model but intelligence is still the core.
Harness Engineering for Self-Improvementlilianweng.github.io
30 June
-
Their words
It becomes insufferable like as a mathematician because you you would basically be like I'm every single time I see one of these I kind of don't know if it's worth my time even if 99 out of 100 of them are right.
27 June
-
korrents.com
When an AI lab builds both a model and a coding harness, the model is optimized for that lab's own harness first.Their words
if an LLM developer also develops a coding harness, it is somewhat safe to assume that their model is optimized for their own harness first
15 June
6 June
27 May
-
Their words
And our harness wasn't very good for like the first like five months of open code. But it was good enough. It was good enough that most people couldn't really tell a difference. And once we won enough share, then we went back and like tried to make our harness like good and smart and optimize and all those things. But uh it was inverted from what everybody else was doing. Everybody else was being like you have to build the smartest harness and that's how you win.
18 May
7 April
-
Their words
and I don't think it's possible for one person to understand everything in depth. Lots of people understand broadly a lot of these ideas but they don't understand sort of everything in in the depth that is actually utilized. That's why there's these, you know, papers with with well over a thousand authors.
2 April
20 March
From one piece Terence Tao – How the world’s top mathematician uses AI 2 beliefs, in the piece's order there
-
Their words
Um I think with within a decade, a lot of things that mathematicians currently do um we we spend a lot of the bulk of our time doing and a lot of stuff we put in our papers today can be done by AI. Um but we will find that that actually wasn't the most important part of what we do.
-
Their words
Uh yeah, so it's it's made papers sort of richer and broader, but not necessarily deeper.
-
Their words
But okay, they shouldn't actually be enacting these ideas. There is a queue of ideas and there's maybe an automated scientist that comes up with ideas based on all the archive papers and GitHub repos and it funnels ideas in or researchers can contribute ideas, but it's a single queue and there is workers that pull items and they try them out.
24 January
-
Lovedaffiliate linkrcmnd.app
Wyatt harness bootsTheir words
I've worn my Wyatt harness boots for years and they're still pristine.
7 January
-
Recommendsaffiliate linkrcmnd.app
Chatter: The Voice in Our Head, Why It Matters, and How to Harness ItTheir words
can lead to imposter phenomenon. In this book, a leading psychologist looks at the science behind our inner voice and new research into how to harness it and improve your physical and mental health.
30 December 2025
23 November 2025
1 November 2025
13 October 2025
-
Their words
Over the past few decades, journal editors and peer-reviewers have increasingly insisted that papers must present large datasets that have been treated using complex statistical methods in order to make even the mildest claims about what caused what.
25 September 2025
4 September 2025
-
korrents.com
The papers claiming Spinosaurus was a diver or a strong swimmer rest on very weak evidence at best.Their words
So Spinosaurus is a super weird and exaggerated version of what was already a kind of super weird group of theropods. So Spinosaurus is properly strange. And then, as you kind of hinted at, super controversial as well, because various papers have claimed it's a diver or a really good swimmer, and I think the evidence for that is very weak at best.
27 August 2025
9 August 2025
-
Their words
One of my favorite papers about prompt injection is Design Patterns for Securing LLM Agents against Prompt Injections
14 June 2025
From one piece Terence Tao: Hardest Problems in Mathematics, Physics & the Future of AI | Lex Fridman Podcast #472 2 beliefs, in the piece's order there
-
Their words
There are certainly math results which could only have been accomplished because there was a human authentication and an AI involved, but it's hard to disentangle credit. I mean, these tools, they do not replicate all the skills needed to do mathematics, but they can replicate some non-trivial percentage of them, 30, 40%, so they can fill in gaps.
-
Their words
And that's a phase shift, because suddenly it makes sense when you write a paper to write it in Lean first, or through a conversation with AI, which is generally on the fly with you, and it becomes natural for journals to accept.
10 June 2025
7 June 2025
-
Their words
Worse, as the latest Apple papers shows, LLMs may well work on your easy test set (like Hanoi with 4 discs) and seduce you into thinking it has built a proper, generalizable solution when it does not.
5 May 2025
-
Their words
I don’t think anyone really believes that at the event horizon you’ll find a firewall. But it did lead to things like the entangled wormholes embroidering a black hole, which was born out of an attempt to address the concerns that AMPS raised. So it did lead to progress.
7 January 2025
24 November 2024
6 November 2024
-
Lovedrcmnd.app
CiteSeerTheir words
Online library of research papers.
18 July 2024
5 June 2024
11 April 2023
-
Their words
Reviewers are less likely to greenlight papers and grants if they’re novel, risky, or interdisciplinary.
9 March 2023
13 December 2022
From one piece The rise and fall of peer review 2 beliefs, in the piece's order there
-
Their words
Scientists have run studies where they deliberately add errors to papers, send them out to reviewers, and simply count how many errors the reviewers catch. Reviewers are pretty awful at this.
-
Their words
In fact, we’ve got knock-down, real-world data that peer review doesn’t work: fraudulent papers get published all the time.
7 December 2022
-
Their words
It terrifies me, because these data repositories are not only a risk to individual user privacy, they’re effectively a surveillance super-weapon.
Why encrypted backup is so importantblog.cryptographyengineering.com
11 July 2021
20 January 2021
-
korrents.com
Pre-grant peer review predicts which research will have impact substantially better than chance does.Their words
I have reviewed before the effectiveness of peer review at figuring out "what's good" in the context of grant awards, finding that peer review as currently practiced does substantially better than chance at predicting future impact, especially for the most impactful of the papers; but at the same time is is far from perfect, leaving plenty of variance unexplained.
13 September 2020
-
Usesrcmnd.app
Adobe Acrobat ReaderTheir words
To read papers, I use Adobe Acrobat Reader and sync them in the cloud. This lets me read, highlight, and sync my papers across devices (work laptop, personal laptop, iPad). Instapaper does the same for online articles.
-
Usesrcmnd.app
InstapaperTheir words
To read papers, I use Adobe Acrobat Reader and sync them in the cloud. This lets me read, highlight, and sync my papers across devices (work laptop, personal laptop, iPad). Instapaper does the same for online articles.
30 June 2020
5 June 2018
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.