There is nothing wrong with promoting perspectives that have political commitments; in fact, intellectual diversity requires this, because some viewpoints include political commitments.
and we're teaching them the worst habits in the world. And so we're giving them 12 years of the worst habits and then even their capability can't overcome that. That it's the super gifted athlete who literally has every bad habit in the book
the best teacher you had, like everybody remembers the one or two teachers who changed their life. It wasn't because they marked up your paper really well. That wasn't what did it for them. It was because you had some block. They saw you and they convinced you you could do so much more.
The idea that you can have autonomous cloud labs running experiments on a loop without human intervention, testing hypotheses, running the experiments, seeing the results, creating a new hypothesis, testing it and running it, and ultimately getting to a conclusion, that is in our sights.
Like the thing that I think is most likely to be sort of the bottleneck in terms of like the AI are really good at verifiable domains but not not at doing the actual thing is just like big experiments. You only get a few tries um well a few is maybe a bit understated but like basically like historically R&D has been driven by doing near frontier scale experiments and that has been pretty important and like actually doing the one big training run where you decide exactly what to include in that.
And so if nobody knows how much an experiment costs, like every time you like grow up a new cell line or every time we make a new a new probe in in the fab, how much does that loop cost? Nobody knows. Therefore, experiments are free. Um, it doesn't cost dollars. It costs media. And media comes from the fridge.
Um but I think there's no uh you know real impediment to making that be a much more automated loop where the model itself decides it's going to explore or maybe with a nudge from some people uh at the various highest level like oh why don't you try some new ideas around model architectures that incorporate this and then it will go run lots of experiments uh see which ones work and then those will get incorporated at a much more rapid rate
I feel like a relevant insight in learning was um recognizing that like who matters more than what. So, like advice to any college student when they're choosing what courses to take. Uh, care a little bit less about your pre-existing interests because they're kind of arbitrary right now. and care a little bit more about whether like the person teaching it is a good educator and someone you resonate with.
One of my favorite lecture series he has given is The Evidence for Modern Physics where he breaks down the experiments that validate some of the weird laws of physics we have, and what it would take to validate even the weirder ones.
In short, whenever we have high-quality evidence that rigorously compares two teaching methods, the research invariably favors strong, direct instruction plus practice
But okay, they shouldn't actually be enacting these ideas. There is a queue of ideas and there's maybe an automated scientist that comes up with ideas based on all the archive papers and GitHub repos and it funnels ideas in or researchers can contribute ideas, but it's a single queue and there is workers that pull items and they try them out.
you're going to run out of things that you're going to do purely in the digital space. At some point you have to go to the universe and you have to ask it questions. Um you have to run an experiment and see what the universe tells you to get back to learn something.
Trials serve two distinct functions: validation — confirming whether a drug works and is safe — and learning, or generating biological data to refine our understanding of a disease, a compound, and the relationship between the two.
it's it's kind of purely a practical and empirical thing that we've observed that by teaching the model principles, getting it to learn from principles, its behavior is more consistent, it's easier to cover edge cases, and the model is more likely to do what people want it to do.
Tariffs, and especially the erratic implementation of them, are teaching our allies and enemies alike that the US is no longer a reliable trading partner.
Try a number of things, right? So, experiment with a number of things. Run a number of small experiments. Once you find something that works, double down on it and then keep doing it until it stops working, which is the step that a lot of people skip. Um, and then once it stops working, go back to the start and try a lot of small experiments again.
And then the top down belief is the thing that sustains you when the experiments contradict you. Because if you just trust the data all the time, well, sometimes you can be doing a correct thing, but there's a bug.
Oulipo's co-founder Raymond Queneau has a book to include here called Exercises in Style (trans. by Barbara Wright), which retells an intentionally banal story 99 times.
But even more important than efficiency, designed experiments can inform about causality, which is very difficult to determine from collected observed data.
I mean in my experience the typical win rate and I hate to use that term for experiments is is often something like 30 to 50%. Like usually you're not actually get like you're trying a bunch of things a lot of hypotheses turn out not to be true. you know, consumer products are very unpredictable like that.
And if I'm starting to see like more and more experiments that are not statistically significant, that may be a signal to me to say, okay, we might have kind of tried to exploit a little bit too far. Like there might not be as much juice to squeeze.
The system matters just as much as any given experiment. Probably even more. Right. I think starting with a growth model so you have an understanding of how your company grows in the first place and which channels you're going to leverage is critical. You need to make sure that you are instrumenting your product in and out otherwise you're going to run experiments and have wonky results.
And a thousand experiments by itself, like if you just did that, but you didn't learn, you didn't make an impact, and that's kind of a waste of time, right? The whole point of setting a goal is that you can have conversations about what would need to be true to actually hit that goal.
The overall workflow is intuitive, especially for those new to formal evaluation processes. The UI guides you through creating datasets, running experiments, and annotating results.
Because some people are simply better at school than other people, any pedagogical strategy, practice, or method that improves the performance of the worst students will also improve the performance of the best students
A more how-to pragmatic version of carving your own path. Includes some personal story of reinvention but more on experimentation that challenges our default scripts of success and ambition.
Accepted practice is that for any given model that is a notable advancement, you're going to do two to 4x compute of the full training run in experiments alone.
This is why you want to work in post-training because the GPU cost for training is lower. So you can make a higher percentage of your training runs YOLO runs.
I really believe that the founder led growth is not being popularized enough that you do not need growth teams until you actually can start running experiments on your user base
And so, what we found was basically that there have been 10 billion trillion habitable zone planets in the universe. And what that means is that those are 10 billion trillion experiments that have been run. And the only way that we're this whole process from a biogenesis to a civilization has occurred is if every one of those experiments failed.
I don’t think that Atlantis existed. I do think it was one of Plato’s many parables talking about putting it in an interesting story as a teaching device in his school.
On the other hand, the manager-free experiments I’m aware of (e.g. holacracy at Medium and GitHub, or “Choose Your Own Work” at Linden Lab) have all been quietly abandoned or outgrown.
Grade harder, grade softer, grade with words rather than numbers, hide the grades entirely—none of it works, because the problem was never the grades themselves. The problem was trying to be both a teacher and a cop at the same time.
What we could we could train people by saying, "Here's the recognition prime decision model. Now, you know, follow this." And that would be useless. Because of uh telling people the strategy isn't going to buy them anything. The intuition part is a reflection of the patterns they've built up through experience. And so there's no shortcut for that.
Though the book perhaps gets a little too math-ey by the end (and this is coming from someone who is literally teaching discrete mathematics to university students at the moment), Dantzig provides a compelling overview of what we mean when we talk about "numbers," and why this is a question that we must repeatedly keep re-visiting.
undergoing some of these experiments was incredibly fun and a great excuse to study a number of textbooks on biochemistry (I liked "Molecular Biology of the Cell"), biology (I liked Campbell's Biology), human nutrition (I liked "Advanced Nutrition and Human Metabolism"), etc.
undergoing some of these experiments was incredibly fun and a great excuse to study a number of textbooks on biochemistry (I liked "Molecular Biology of the Cell"), biology (I liked Campbell's Biology), human nutrition (I liked "Advanced Nutrition and Human Metabolism"), etc.
undergoing some of these experiments was incredibly fun and a great excuse to study a number of textbooks on biochemistry (I liked "Molecular Biology of the Cell"), biology (I liked Campbell's Biology), human nutrition (I liked "Advanced Nutrition and Human Metabolism"), etc.
Compassionate Inquiry is an in depth teaching and distillation of the approach I have developed to working with human beings beset by personal issues, health problems that need gently guided exploration, mental health challenges, addictions, relationship difficulties and, above all, an unhealthy relationship to their own selves.
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com.
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.