Related posts
François Chollet
GitHub
deep-learning-models — Keras code and weights files for popular deep learning models.
Keras weights
The subject this post names, from the same vocabulary
the directory files beliefs under, and the word it uses that this site has seen
least often elsewhere. Posts are matched on those words alone —
nothing here is a summary of this one.
Sources
All
Writing
x.com Blog Korrents Mastodon Bluesky Newsletter GitHub
Top people
Dan Cederholm
François Chollet
Miles Brundage
Jan Leike
Nathan Lambert
Andrej Karpathy
Olivier Blanchard
Ismail Ghallou
Daniel Lemire
Andrew Gelman
Sebastian Raschka
Ben Carlson
Showing
Profile →
Show everything
Hiding
Show them again
Show them again
Further back ↓
Hiding
Show them again
19 September
Economist; former chief economist of the International Monetary Fund.
Ratings largely determine spreads. Yet, their construction remains largely a black box. Our paper assesses the relative importance of debt, deficits and country specific effects in the determination of these ratings. The evidence is at odds with a simple model of default: more importance of debt, less importance of forecast deficits, limited role of r-g, surprisingly large role of country effects. We are agnostic. The model may be too simple in some fundamental ways. Or the rating agencies may be using incorrect weights. Or a mix of the two... We hope the paper triggers a useful discussion. Quoting @int_mon_econ Super interesting! "Ratings, Debt, and Deficits: An Exploration" by Olivier J Blanchard, Daniel Leigh, and Prachi Mishra. "We look at the effects of debt and primary fiscal balances on sovereign credit ratings through the lens of a simple model. We find that the ratings differ from the implications… Related
17 September
Creator of the Keras deep-learning library and the ARC-AGI benchmark.
14 September
Moroccan senior frontend developer writing at smakosh.com; builds side projects and client work under Smakosh LLC.
Start by open-sourcing your models and sharing the weights then Quoting @sama The world deserves confidence that American companies developing increasingly capable AI will act responsibly, especially as the trajectory of progress has steepened. Every frontier lab must deliver on this, and there is no reason any of us should come to work if we cannot. We welcome a federal fra… Related
11 September
Computer science professor who works on fast data processing; co-author of the simdjson parser and a weekly blogger about software performance since 2004.
What should be obvious is that inference, running a large language model, is the part that has to be cheap. Most of our hardware was not designed for that. LLM inference is limited by bandwidth. You do a great many matrix-vector multiplications against a huge weight matrix, and you barely reuse those weights. It is closer to streaming a video than to running a simulation or drawing a scene in a game. Graphics processors from Nvidia and others were built for something else. They pileed compute first, then bolted on high-bandwidth memory. It is expensive. An obvious answer is to invert the desi… LLMs Related
9 September
Creator of the Keras deep-learning library and the ARC-AGI benchmark.
ZeroModels: 100+ model families with pretrained weights, implemented in Keras 3 and ready to run with any backend. Between this and KerasHub, that's a very extensive range of foundation models. imvision12.github.io
Keras Related
7 September
Statistician at Columbia; writes Statistical Modeling, Causal Inference, and Social Science, and wrote Bayesian Data Analysis.
6 September
AI research engineer working on large language models. He writes the Ahead of AI newsletter and is the author of Build a Large Language Model (From Scratch).
Reasoning from scratch round 2: In this video, I cover the text generation process in LLMs and KV caching (to prepare the base model before adding reasoning techniques in the upcoming ones). 00:00 Introduction and reasoning model demo 01:55 How to work through the book 05:00 Chapter 2 overview 08:25 Checking PyTorch and hardware support 10:26 Apple silicon and MPS caveats 15:00 Cloud GPU options 16:08 Tokens and tokenization 18:20 Qwen3 and the Reasoning From Scratch package 23:05 Encoding and decoding text 26:24 Downloading weights and selecting a device 31:01 Loading the pretrained Qwen3 mo… LLMs Related
Director of institutional asset management at Ritholtz Wealth Management. Writes the A Wealth of Common Sense blog and co-hosts the Animal Spirits podcast with Michael Batnick.
2 September
Investor and writer. Previously a partner at Andreessen Horowitz and a product leader at Twitter, Facebook, Snap and Microsoft. Writes essays and memos at sriramk.com.
great momentum from @finkd and @alexandr_wang. particularly excited by their upcoming open weights launch - we need more American open weight models at the frontier. America Related
1 September
Writes Hyperdimensional, a newsletter on AI policy and governance. A White House AI policy adviser in 2025; joined OpenAI on 6 July 2026 to lead its Strategic Futures team.
19 August
Web designer; founder of SimpleBits, co-founder of Dribbble, and author of Bulletproof Web Design and Handcrafted CSS.
Mastodon Hubano 2.0 now available! I added two additional weights, some new glyphs and symbols, and lots of fixes and tweaks. Free update for those who purchased Hubano as well as Secret Club™ members. simplebits.com
Related
Bluesky Hubano 2.0 now available! I added two additional weights, some new glyphs and symbols, and lots of fixes and tweaks. Free update for those who purchased Hubano as well as Secret Club™ members. simplebits.com
Related
14 August
Machine-learning researcher on open language models; writes the Interconnects newsletter and the RLHF Book, after leading post-training at Ai2.
Their words
At the end of the day, this type of safety barely matters when true open-weights are coming.
Show the whole quote
interconnects.ai
12 August
Web designer; founder of SimpleBits, co-founder of Dribbble, and author of Bulletproof Web Design and Handcrafted CSS.
Mastodon A free pre-release of my Hubano 2 typeface is available now for Secret Club members. New weights, fixes, and additional glyphs! simplebits.com
Related
Bluesky A free pre-release of my Hubano 2 typeface is available now for Secret Club members. New weights, fixes, and additional glyphs! simplebits.com
Related
9 August
Web designer; founder of SimpleBits, co-founder of Dribbble, and author of Bulletproof Web Design and Handcrafted CSS.
Mastodon Some glyphs in the new weights coming very soon to Hubano, my typeface inspired by Cuban cigar box lettering. Related
Bluesky Some glyphs in the new weights coming very soon to Hubano, my typeface inspired by Cuban cigar box lettering. Related
26 July
Full-stack developer, creator of Tailwind CSS; writes at adamwathan.me and makes courses and screencasts for web developers.
📊 Do you lift weights at least twice a week most weeks? Related
22 July
Security engineer; founder of Matasano Security and Latacora.
OpenAI
Their words
I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes.
Show the whole quote
@tqbf on X x.com
20 July
Machine-learning researcher on open language models; writes the Interconnects newsletter and the RLHF Book, after leading post-training at Ai2.
Physician and neuroscientist at the University of Washington; researches brain health, metabolism and athletic performance.
Their words
Whereas aerobic exercise on that kind of intensity spectrum we just talked about seems to particularly benefit the gray matter which includes the hippocampus. Resistance training seems to particularly benefit the white matter.
Show the whole quote
youtube.com
15 July
Web designer; founder of SimpleBits, co-founder of Dribbble, and author of Bulletproof Web Design and Handcrafted CSS.
Mastodon Kyroh is my latest Danmade typeface, inspired by a vintage Egyptian hotel luggage tag. Comes in 3 weights with a no-fuss license—or join my Secret Club™ and get all of my fonts + other perks. simplebits.shop
Related
Bluesky Kyroh is my latest Danmade typeface, inspired by a vintage Egyptian hotel luggage tag. Comes in 3 weights with a no-fuss license—or join my Secret Club™ and get all of my fonts + other perks. simplebits.shop
Related
7 July
Web designer; founder of SimpleBits, co-founder of Dribbble, and author of Bulletproof Web Design and Handcrafted CSS.
Mastodon My latest typeface, Kyroh, is out today! Three weights of a Danmade, imperfect display sans based on an old Egyptian hotel luggage tag. Get it on its own or along with all my fonts by joining my Secret Club ✨ simplebits.shop
Related
Bluesky My latest typeface, Kyroh, is out today! Three weights of a Danmade, imperfect display sans based on an old Egyptian hotel luggage tag. Get it on its own or along with all my fonts by joining my Secret Club ✨ simplebits.shop
Related
5 January
Co-founder of Modem and a founding engineer at Sentry, where he went on to be VP of Engineering. Co-author of Third-party JavaScript, and co-host of the State of Agentic Coding podcast with Armin Ronacher.
17 October 2025
Founding member of OpenAI and former director of AI at Tesla; creator of nanoGPT and the term "vibe coding".
From one piece
Andrej Karpathy — “We’re summoning ghosts, not building animals”
2 beliefs, in the piece's order there
Their words
anything that's in the weights it's kind of like a hazy recollection of what you read a year ago anything that you give it as a context uh at test time is directly in the working memory
Show the whole quote
youtube.com
Their words
These models don't really have this distillation phase um of taking what happened, analyzing it, obsessively thinking through it, um basically doing some kind of a synthetic data generation process and distilling it back back into the weights
Show the whole quote
youtube.com
11 March 2025
Independent AI policy researcher; led policy research at OpenAI from 2018 to 2024, latterly as senior adviser for AGI readiness.
10 January 2025
Independent AI policy researcher; led policy research at OpenAI from 2018 to 2024, latterly as senior adviser for AGI readiness.
20 December 2024
Independent AI policy researcher; led policy research at OpenAI from 2018 to 2024, latterly as senior adviser for AGI readiness.
8 November 2024
Alignment researcher at Anthropic; co-led OpenAI's superalignment team until 2024, and writes Musings on the Alignment Problem.
19 June 2024
Co-founder and CEO of Perplexity, an AI answer engine; previously a research scientist at OpenAI and a PhD student at UC Berkeley.
Their words
It’s less about access to a model’s weights, it’s more access to compute that is putting the world in more concentration of power and few individuals. Because not everyone’s going to be able to afford this much amount of compute to answer the hardest questions.
Show the whole quote
youtube.com
18 March 2024
Co-founder and CEO of OpenAI; previously president of Y Combinator.
Their words
We want to put increasingly powerful tools in the hands of people for free and get them to use them. I think that kind of open is really important to our mission. I think if you give people great tools and teach them to use them or don’t even teach them, they’ll figure it out, and let them go build an incredible future for each other with that, that’s a big deal.
Show the whole quote
youtube.com
12 February 2024
Creator of the Keras deep-learning library and the ARC-AGI benchmark.
1 November 2023
Creator of the Keras deep-learning library and the ARC-AGI benchmark.
15 March 2023
Machine-learning researcher; has written the Lil'Log survey posts on how a model technique works since 2017, and worked at OpenAI from 2018 to 2024, latterly leading its safety systems team.
Prompt Engineering , also known as In-Context Prompting, refers to methods for how to communicate with LLM to steer its behavior for desired outcomes without updating the model weights. It is an empirical science and the effect of prompt…
LLMs Related
2 February 2022
Co-founded GiveWell and led Open Philanthropy; writes Cold Takes, on the most important century and how to think about it.
9 March 2018
Associate Professor at MIT CSAIL, leading the Programming Systems Group; works on programming systems that manipulate uncertainty, from neural networks to unreliable hardware.
Their words
The winning tickets we find have won the initialization lottery: their connections have initial weights that make training particularly effective.
Show the whole quote
arxiv.org
Chief AI Scientist at Databricks, after its acquisition of MosaicML, where he was a founding team member; known for the Lottery Ticket Hypothesis from his MIT PhD (ICLR 2019 Best Paper).
Their words
The winning tickets we find have won the initialization lottery: their connections have initial weights that make training particularly effective.
Show the whole quote
arxiv.org
2 February 2018
Creator of the Keras deep-learning library and the ARC-AGI benchmark.
Nothing matches. Show everything
What is a korrent?
A korrent is a belief a person has stated in their own words: one
sentence stating the claim, backed by a quote and a source, kept at
korrents.com .
Under a name here, the quoted block is what they actually said.
The korrent beneath it is the claim those words support, in
korrents' wording — tap it to see the record, its source, and who
else holds it.
Nobody here wrote their own korrents. They are compiled from public
statements, and a person can change their mind, which is recorded too.
Got it
About the English under a post
Some people here publish in a language other than English. Where they
do, this site shows a machine translation beneath the post, in
this typeface — the site's own, not theirs.
The post itself is never changed, moved or hidden: what is set in the
serif above is exactly what the person published, and it is what to
quote them on. A translation can be wrong in ways that matter,
especially about tone.
Only the post's own words are translated. A quoted post, a linked
article and a belief on korrents.com
are left in their original language.
Got it