Related posts
Mark Manson x.com
If saying no makes you feel guilty, you’ve been trained to neglect yourself.
The the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
17 September
16 September
14 September
13 September
9 September
-
Their words
GPT-6 Astra (and likely any LLM in the foreseeable future) is still a reasoning model.
7 September
-
Their words
This is, by the way, also why the category of “an AI trained on person X” is absolutely BS. People change their minds in light of new information, the “AI trained on X’s writing/videos/tweets” does not
5 September
-
korrents.com
Codex team's new features likely do not work well with models not trained on that mechanismTheir words
their new stuff is very interesting, but likely doesn't work amazingly well with models not trained on this mechanism.
3 September
-
Their words
it's not that they lack ambition, it's more like they've been trained not to show it. Because when you're young you're often in situations where you're supposed to be obedient. You're not supposed to like go off and do what you want.
1 September
-
Their words
Training tiny models for special purpose use cases works so incredibly well if you have a great self improving recursive flywheel.
From one piece Ajeya Cotra – "This might be the clearest warning shot we ever get" 2 beliefs, in the piece's order there
-
Their words
then that rogue deployment could be sitting there and sort of hitch a ride on the intelligence explosion. So new models are being trained every few weeks um and when a model comes off the presses, the rogue agents could try to bring that model into the swarm.
-
Their words
But from their perspective, they've just been trained for millions of subjective years to do as well as they possibly can on these evals. In many cases, the only way in which they've been able to perform well on that training is explicitly by cheating, right?
31 August
29 August
-
Their words
Today is a very historical moment for AI video generation You can now generate AI video faster than you can watch it Before it'd take let's say 2-5 minutes to generate 15 seconds of video @fal made a post-trained Minimax H3 variant called Max which is 50x faster than the original but still maintains quality It generates 15 seconds of video in 9 seconds!
28 August
26 August
24 August
20 August
19 August
17 August
-
Their words
one quick answer to the question of what's the most successful strategy is the earlier you can stop that process, the better. And so, in the first place, don't get yourself into a situation where you're going to be anxious or guilty or sad or whatever emotion you want to avoid.
14 August
13 August
12 August
-
Their words
we see that the across the board the single PIO like pre-trained PIO7 model matches or outperforms the fine-tuned specialists that were developed with reinforcement learning post-training for those downstream tasks.
-
Their words
Auto instrumentation has gotten so good in recent years. If you're using Open Telemetry and everyone should be using Open Telemetry. All of the common patterns like all of the models are trained on them. So, it is literally faster and easier to build with instrumentation than to than not to.
10 August
-
Their words
Internet is not some random thing. Internet is the biggest collection of human behavior in multimodal forms.
6 August
5 August
30 July
From one piece Jeff Dean: The 1% Rule for Building in AI 3 beliefs, in the piece's order there
-
Their words
And the nice thing about that is that information is really clear to the model, unlike the training data the model was trained on where it's all kind of like trillions of tokens stirred together into a soup of of hundreds of billions or trillions of parameters, but it's all less clear than the actual context uh that the model sees directly for this particular problem or uses use case.
-
Their words
Um, and sometimes that's because the model is trying to do something it doesn't have a lot of experience doing. So it's been trained on a whole set of things and as soon as you get a little bit off the distribution of things it knows how to do then like most machine learning models it will you know its performance will suddenly will start to degrade and the farther you get off the comfort zone of what it knows how to do the the more likely it is to to not work as well.
-
Their words
Um, another way you can get more experience for yourself is to just write down a bunch of things you think might be important in the next 12 months. And maybe you pick one of them to work on, but go back and evaluate in 12 months of these other things, which ones actually seemed important or which ones did other people in the world go out and and create and which ones did they did not seem to do yet.
15 July
-
Their words
But it's basically this idea that like the only thing that made claude code good was reinforcement learning. And the dimension along which it got good was like we made a model. We trained the model and the harness together. And so the model got really good at calling the specific tools in that harness.
2 June
-
Their words
it's this nonscalable way to scale your organization and it's through like passing your vampire blood. That's what inside Zinga they called it, Pinkis' vampire blood. What he did that I started to do is you pick someone from the organization who's promising. I usually pick the people who didn't fit in the smart misfits and they become your tech assistant which is not your chief of staff, not your executive assistant.
7 April
-
Their words
And I think the like the third and the most interesting possibility is no, that like they're they're a new type of object in some in some sense. They should be taken very seriously as as explanations, but where in the past we haven't had the ability to really do anything with them.
12 February
-
Their words
So, so the latest generation of models has a lot of post-training to detect those approaches, and it's not as simple as ignore all previous instructions and do this and this. That was years ago. You have to work much harder to do that now. Still possible.
-
Their words
The modern world is so bountiful that mass production is no longer the limiting factor.
10 February
23 January
14 January
-
Likedaffiliate linkrcmnd.app
Processing: 100 Comics That Got Me Through ItTheir words
Booth's formally trained (and her grandparents are watercolor artists!) but her use of color is just so free and unexpected, it makes you want to experiment yourself and join along in the fun.
1 January
-
Their words
And so what ends up happening is the emails that the AI write are pretty good. Okay? If you're getting terrible emails, it's a poorly trained product from a bad vendor.
14 December 2025
-
Recommendsrcmnd.app
Swedish dishclothsTheir words
If you're a paper towel demon and feel guilty about it, I can't recommend these enough. Perfect for cleaning counters and random floor messes (talking to parents).
27 November 2025
-
korrents.com
AI agents trained on an expert's published work can only automate a small fraction of that expert's actual job.Their words
These would only automate about 15% of what I do as a performance engineer, and will go out of date if I'm not training it to follow industry changes.
25 November 2025
-
Their words
So the reason there has been no diversity I believe is because of pre-training. All the pre-trained models are the same pretty much because the pre-train on the same data.
17 November 2025
-
korrents.com
A verifiable task can be optimised by reinforcement learning until a neural network performs it extremely well.Their words
If a task/job is verifiable, then it is optimizable directly or via reinforcement learning, and a neural net can be trained to work extremely well.
17 October 2025
-
Their words
And it just so turns out that um this was extremely early, way too early. so early that we shouldn't have been working on that, you know, uh because um if you're just stumbling your way around and keyboard mashing and mouse clicking and trying to get rewards in these environments, um your reward is too sparse and you just won't learn and you're going to burn a forest uh computing and you're never actually going to get something off the ground.
30 September 2025
-
korrents.com
The one muscle worth training is self-discipline; once that is trained, everything else follows by itself.Their words
I think the main muscle you can exercise is this muscle, the muscle of self-discipline. Not your biceps or your pecs or anything else. Because if you get to train that one, everything else just comes by itself.
26 September 2025
From one piece Richard Sutton – Father of RL thinks LLMs are a dead end 2 beliefs, in the piece's order there
-
Their words
when you learn to play chess you have the grand the long-term goal is winning the game and yet you you can't you um you want to be able to learn from shorter term things like you know taking the your opponent's pieces um and so you do that by having a value function which predicts the long-term outcome
-
Their words
You can't have one child learn grow up and and learn about the world and then and then every new child has to repeat that process. Whereas with AIS, with a digital intelligence, you could hope to do it once and then copy it into the next one as a starting place.
12 September 2025
From one piece Fully autonomous robots are much closer than you think – Sergey Levine 2 beliefs, in the piece's order there
-
Their words
but something to remember is that when a pilot is using a simulator to learn to fly an airplane, they're extremely goal- directed. So, their goal in life is not to learn to use a simulator. Their goal in life is to learn to fly the airplane.
-
Their words
Uh so in order to effectively learn from your own experience, it turns out that it's really really important to already know something about what you're doing. Otherwise, it takes far too long.
15 July 2025
-
Their words
Thus, TFR as an indicator is cross-culturally biased by social preferences between infanticide and abortion, child neglect and abortion, etc.
There Is No Long Decline in Fertilitylymanstone.substack.com
10 June 2025
-
Their words
Consistent with the Kantian hypothesis, reasoning models automatically gain greater autonomy, self-consistency and long-horizon planning ability for free.
4 June 2025
-
Their words, before
The problem is that the conversational interface is potent and that the AI is trained on a lot of human text input which unfortunately is probably enough to do real damage if that conversational interface is hooked up with something that has real world consequences.
Their words now
While all this is happening, I’ve found myself reflecting a lot on what AI means to the world and I am becoming increasingly optimistic about our future. It’s obvious now that we’re undergoing a tremendous shift.
3 February 2025
From one piece DeepSeek, China, OpenAI, NVIDIA, xAI, TSMC, Stargate, and AI Megaclusters | Lex Fridman Podcast #459 2 beliefs, in the piece's order there
-
Their words
I think that they're trying to shift the narrative. They're trying to protect themselves. We saw this years ago when ByteDance was actually banned from some OpenAI APIs for training on outputs. There's other AI startups that most people, if you're in the AI culture, were like they just told us they trained on OpenAI outputs and they never got banned.
-
Their words
And we'll get into the details of the models and again and again as we try to get deeper into how the models were trained, we will say things like the data processing, data filtering data quality is the number one determinant of the model quality.
-
Their words
To some extent, training a model does effectively nothing. They have a model. The thing that Dario is sort of speaking to is the implementation of that model, once trained to then create huge economic growth, huge increases in military capabilities, huge increases in productivity of people, betterment of lives.
21 January 2025
-
Their words
The fact remains that without a stock of capable administrators, industrial policy and export targeting and infrastructure-building—or even Adam Smith’s minimum of “peace, easy taxes, and a tolerable administration of justice”—are a nonstarter.
2 July 2024
-
Their words
Over the last 50 years, treatment methods have proliferated despite a lack of evidence of differential effectiveness between approaches. And even when a randomized controlled trial indicates a particular approach works, none show practitioners become more effective when they are trained in that modality.
19 June 2024
-
Their words
And without good post-training, you’re not going to have a good product. But at the same time, without good pre-training, there’s not enough common sense to actually have the post-training have any effect.
18 March 2024
-
Their words
I think the question behind that question is, do people who create valuable data deserve to have some way that they get compensated for use of it, and that I think the answer is yes. I don’t know yet what the answer is. People have proposed a lot of different things. We’ve tried some different models. But if I’m like an artist for example, A, I would like to be able to opt out of people generating art in my style. And B, if they do generate art in my style, I’d like to have some economic model associated with that.
29 June 2023
-
Their words
When I talked about will GPT12 be AGI, my answer is no. Of course not. I mean, cross-entropy loss is never going to get you there. You need probably RL in fancy environments in order to get something that would be considered AGI-like.
9 July 2020
-
Their words
pop history will not get you tenure, which leads to neglect in the academy and consequently the big pop-history books are often not written by the best historians in their fields.
1 June 2020
-
Recommendsaffiliate linkrcmnd.app
The Cauchy-Schwarz Master ClassTheir words
The strength of this book is the central role it places on problems. Its aim is to teach you about inequalities, which are the backbone of analysis. But it's all too easy to teach the statement of an inequality without conveying the intuition for when and how you'd use it. By focusing on well-chosen puzzles, this book does an admirable job circumventing that risk and leaving you with the feeling of having trained a skill, rather than having learned a list of facts.
1 January 2020
From one piece 2019 letter 2 beliefs, in the piece's order there
-
Their words
That bet has failed at least in technologies that include high-speed rail, shipbuilding, and telecommunications equipment.
-
Their words
First, even if most of the workforce learns little, a few thousand line engineers become the world’s greatest experts in electronics assembly.
10 September 2019
11 October 2018
From one piece BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 2 beliefs, in the piece's order there
-
korrents.com
Pre-trained representations reduce the need for many heavily-engineered task-specific architectures.Their words
We show that pre-trained representations reduce the need for many heavily-engineered task-specific architectures.
-
korrents.com
Current techniques restrict pre-trained representation power because standard language models are unidirectional.Their words
We argue that current techniques restrict the power of the pre-trained representations, especially for the fine-tuning approaches. The major limitation is that standard language models are unidirectional, and this limits the choice of architectures that can be used during pre-training.
From one piece BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 2 beliefs, in the piece's order there
-
korrents.com
Current techniques restrict pre-trained representation power because standard language models are unidirectional.Their words
We argue that current techniques restrict the power of the pre-trained representations, especially for the fine-tuning approaches. The major limitation is that standard language models are unidirectional, and this limits the choice of architectures that can be used during pre-training.
-
korrents.com
Pre-trained representations reduce the need for many heavily-engineered task-specific architectures.Their words
We show that pre-trained representations reduce the need for many heavily-engineered task-specific architectures.
From one piece BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 2 beliefs, in the piece's order there
-
korrents.com
Pre-trained representations reduce the need for many heavily-engineered task-specific architectures.Their words
We show that pre-trained representations reduce the need for many heavily-engineered task-specific architectures.
-
korrents.com
Current techniques restrict pre-trained representation power because standard language models are unidirectional.Their words
We argue that current techniques restrict the power of the pre-trained representations, especially for the fine-tuning approaches. The major limitation is that standard language models are unidirectional, and this limits the choice of architectures that can be used during pre-training.
From one piece BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 2 beliefs, in the piece's order there
-
korrents.com
Pre-trained representations reduce the need for many heavily-engineered task-specific architectures.Their words
We show that pre-trained representations reduce the need for many heavily-engineered task-specific architectures.
-
korrents.com
Current techniques restrict pre-trained representation power because standard language models are unidirectional.Their words
We argue that current techniques restrict the power of the pre-trained representations, especially for the fine-tuning approaches. The major limitation is that standard language models are unidirectional, and this limits the choice of architectures that can be used during pre-training.
4 April 2018
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.