Okay, if I'm not going to read this code, how do I know it's going to perform within boundaries of the last code that I generated? That is conformance testing.
Despite these strides, we argue that current AI4Math systems still largely operate as solvers, excelling at isolated, well-defined proof generation rather than as researchers capable of expanding the boundaries of mathematical knowledge.
I think the reason why it doesn't show up as so much better on the benchmarks is because the benchmarks are being presented, the benchmark results are being presented in the wrong way. They're not controlling for the amount of test time compute that is being used on that benchmark question.
I don't think that human preference is just an equation that people can just like package nicely and like, "Hey, ask people like which which of these two answers people would prefer?" It's very very personal, very culturally dependent, geographically dependent, age dependent.
I think the most interesting question is what can the models we have do right now? And so the only thing I care about today is what can Claude Opus 4.6 do that we haven't figured out yet. And I think it will take us 6 months to even start exploring the boundaries of that.
a worry I have is that the growth rate could be like 50% in Silicon Valley and, you know, parts of the world that are kind of socially connected to Silicon Valley and, you know, not that much faster than its current pace elsewhere. And I think that'd be a pretty messed up world.
And so the main thing that I felt so excited about and this is something I tell my team all the time is we need to dissolve the boundaries of these traditional roles.
The relative absence of pre-colonial states meant that the newly independent African states were largely European inventions; modern Kenya, for instance, inherits the boundaries of British East Africa, which was created by pulling together the Kikuyu, Luo, Luhya, Kalenjin, and Kamba ethnic groups.
For the vast majority of human existence, we’ve been nomadic and we’ve done these wider or tighter nomadic circles, depending on the geographic region, but they’d move. So once humans figured out how to stay in a place, that’s the initial trigger to what would become civilization.
So what’s fundamental, the way we talk about it in current physics is not actually fundamental, it’s the boundaries of what we can observe in our universe, what we can see with our technology.
It just kind of vanished, because what happens, I’m creating all these artificial event boundaries. I’m losing all this executive function every time I switch, I’m getting a few seconds slower and I’m catching up mentally to what’s happening.
That’s different than something being emergent or non-fundamental, but also real. This table is real, even though I know it’s made of atoms, that doesn’t remove the realness from the table. I think that consciousness and free will and things like that are just as real in tables and chairs.
In traditional classroom settings, the development of skills often gets overlooked in favor of facts. Facts in the form of conjugation tables, word lists, grammar explanations, the names of verb tenses, and so on. All of these things are related to language, but it’s important to realize that knowing these things has nothing to do with speaking and using a language productively.
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com.
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.