Related posts
Lilian Weng Blog
generalizationdatasetparameters
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
16 September
15 September
12 September
11 September
8 September
6 September
-
korrents.com
The fundamental challenge of alignment is generalization: holding values in situations the training never covered.Their words
The fundamental challenge of AI alignment is generalization.
2 September
24 August
15 August
12 August
-
Their words
we find that the performance on held out tasks decreases dramatically. Whereas if we um just take out a random 20% of the data that's less diverse than the most diverse subset, the performance um only decreases a little bit. And so this suggests that actually having really diverse data plays an important role in enabling it to generalize to new tasks.
11 August
-
Likedrcmnd.app
Multi-Paradigm Design for C++Their words
Multi-Paradigm Design for C++ by James O. Coplien. Twenty years after reading this, the core techniques are still with me. These are commonality analysis, where the purpose is to identify families of systems, and variability analysis, which focuses on capturing the domain parameters that vary. In essence: the foundation of great software design.
9 August
-
korrents.com
Ireland's colonial wars helped set the template Britain later applied elsewhere in its empireTheir words
The brutal war and its aftermath set some of the parameters for what Britain did later in other parts of the world.
30 July
From one piece Jeff Dean: The 1% Rule for Building in AI 2 beliefs, in the piece's order there
-
Their words
And the nice thing about that is that information is really clear to the model, unlike the training data the model was trained on where it's all kind of like trillions of tokens stirred together into a soup of of hundreds of billions or trillions of parameters, but it's all less clear than the actual context uh that the model sees directly for this particular problem or uses use case.
-
Their words
And often you can actually make the model work better and succeed at that kind of problem by not just adjusting the model parameters which is hard to do from the outside but from you know creating better guidelines for the model you know writing skills for the model to know how to use different tools that would be incredibly useful for solving this particular class of problem.
22 July
-
korrents.com
The US warming trend since 1970 is consistent across raw, adjusted, satellite, and surface temperature datasetsTheir words
The overall US warming trend (around 0.5F per decade since 1970 in max, min, and mean temperatures) is robust in every dataset, raw or adjusted, satellite or surface.
24 June
10 June
-
Their words
it's not just like there's some factory somewhere that you can pay to produce the the data like you actually need to invent new novel scientific approaches
16 April
13 April
13 February
From one piece Dario Amodei — “We are near the end of the exponential” 2 beliefs, in the piece's order there
-
Their words
the goal is not to teach the model every possible skill within RL just as we don't do that within pre-training, right? Within pre-training, we're not trying to expose the model to, you know, every every possible you know, way that words could be put together, right? You know, we're it's it's rather that the model trains on a lot of things and then and then it reaches generalization across pre-training, right?
-
korrents.com
Models already generalise substantially from tasks that can be verified to tasks that cannot.Their words
We already see substantial generalization from things that that verify to things that don't verify. We're already seeing that.
1 January
20 December 2025
1 December 2025
-
korrents.com
The share of American teenagers who say that life often feels meaningless surged as Gen Z entered adolescence.Their words
But as soon as Gen Z entered the dataset, around 2013, meaninglessness surged.
28 November 2025
23 November 2025
17 October 2025
From one piece Andrej Karpathy — “We’re summoning ghosts, not building animals” 2 beliefs, in the piece's order there
-
Their words
Um so I almost feel like because the internet is so terrible, we actually have to sort of like build really big models to compress all that. Uh most of that compression is memory work instead of like cognitive work. But what we really want is the cognitive part actually delete the memory
-
Their words
the reason that I think this is kind of tricky is quite subtle. And it's the fact that anytime you use an LLM to assign a reward, those LLMs are giant things with billions of parameters and they're gameable.
12 September 2025
-
Their words
Yeah. So there's a subtlety here. Emerging capabilities don't just come from the fact that internet data has a lot of stuff in it. They also come from the fact that generalization once it reaches a certain level becomes compositional.
31 August 2024
From one piece Why A.I. Isn't Going to Make Art 2 beliefs, in the piece's order there
-
korrents.com
Art is what results from making an enormous number of choices, and a prompt contains almost none of them.Their words
But let me offer a generalization: art is something that results from making a lot of choices.
-
korrents.com
Writing that deserves a reader's attention is always the product of effort by the person who wrote it.Their words
Let me offer another generalization: any writing that deserves your attention as a reader is the result of effort expended by the person who wrote it.
11 June 2024
23 May 2024
-
Likedtheir ownrcmnd.app
Decent DE1Their words
In the early days, the DE1 was considered "unforgiving" by many. I didn't mind, as the DE1's ability to control shot parameters more than made up for its early challenges.
20 December 2023
15 November 2023
-
korrents.com
In instruction tuning, a smaller set of high-quality training examples outperforms a larger, noisier dataset.Their words
training on a small set of high-quality data outperforms instruction-tuning on larger, noisier data.
9 October 2023
24 June 2023
-
Their words
At the same time, in order to reduce the probability of someone intentionally or unintentionally bringing about a rogue AI, we need to increase governance and we should consider limiting access to the large-scale generalist AI systems that could be weaponized, which would mean that the code and neural net parameters would not be shared in open-source and some of the important engineering tricks to make them work would not be shared either.
20 July 2022
2 April 2021
-
Their words
Having done the sums before hitting the keys, I was a skeptic. It just didn’t work, even with optimistic assumptions around the parameters for the aircraft and its batteries.
5 November 2019
-
Their words
We argue that solely measuring skill at any given task falls short of measuring intelligence, because skill is heavily modulated by prior knowledge and experience: unlimited priors or unlimited training data allow experimenters to "buy" arbitrary levels of skills for a system, in a way that masks the system's own generalization power.
22 March 2019
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.