Related posts
Sebastian Raschka Newsletter
The subject this post names, from the same vocabulary the directory files beliefs under, and the words it uses that this site has seen least often elsewhere. Posts are matched on those words alone — nothing here is a summary of this one.
Top people
Showing Profile →
Hiding
Hiding
20 September
15 September
13 September
9 September
-
Their words
So, in short, we can say that using looped transformers can improve model quality at a fixed compute budget if the model is large enough.
8 September
-
korrents.com
Hybrid and sparse-attention model architectures will become more widely adopted as the ecosystem catches up.Their words
we expect similar architectures to become more popular and the ecosystem to fix integrations by the time Qwen4 drops.
7 August
30 July
-
Their words
Um but I think there's no uh you know real impediment to making that be a much more automated loop where the model itself decides it's going to explore or maybe with a nudge from some people uh at the various highest level like oh why don't you try some new ideas around model architectures that incorporate this and then it will go run lots of experiments uh see which ones work and then those will get incorporated at a much more rapid rate
28 July
-
korrents.com
Cheaper software has increased the number of software engineers a company needs, not reduced it.Their words
At boom we need far more software engineers in a postAI world than we need in a pre-AI world. Why? Because the cost of software development has dropped. anybody including hardware engineers can now become a coder and we need software engineers to make sure the architectures are right and make sense and are coherent.
16 May
15 April
-
korrents.com
Nvidia does not run rival chip architectures in parallel because it simulates them and they come out provably worse.Their words
Yeah, we we could do all of those things. Um it's just not better. And we simulate it all. they're in our simulator provably worse and so we wouldn't do it.
10 April
22 March
25 February
30 December 2025
11 September 2025
4 September 2025
-
korrents.com
Local-first sync architectures are a natural fit for AI agents, since both are inherently multiplayer.Their words
Agentic and sync are natural complements because local-first is multiplayer by default, and agents are another sort of "player" in addition to human collaborators.
18 August 2025
-
Their words
many of the breakthroughs that we've had that have enabled us to to deliver such quality and cost savings and more have come through novel agent architectures and and really going down a click or two in in the stack to innovate at at lower levels of the technology stack.
5 June 2025
-
Their words
the text embeddings across LLMs appear to largely converge on a "universal geometry" despite differing architectures, parameter counts, and training sets.
16 March 2025
Nothing matches.
What is a korrent?
A korrent is a belief a person has stated in their own words: one sentence stating the claim, backed by a quote and a source, kept at korrents.com.
Under a name here, the quoted block is what they actually said. The korrent beneath it is the claim those words support, in korrents' wording — tap it to see the record, its source, and who else holds it.
Nobody here wrote their own korrents. They are compiled from public statements, and a person can change their mind, which is recorded too.
About the English under a post
Some people here publish in a language other than English. Where they do, this site shows a machine translation beneath the post, in this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the serif above is exactly what the person published, and it is what to quote them on. A translation can be wrong in ways that matter, especially about tone.
Only the post's own words are translated. A quoted post, a linked article and a belief on korrents.com are left in their original language.