AI builders
The people building the systems, in their own words rather than their employers'.
Filter & sortAll sources · condensed
- Boris Cherny bcherny.github.io — My blog
- François Chollet Inside you there are two wolves
- Boris Cherny Projects have changed not only how I interact with Claude but how I code. I stopped managing sessions. I just send thoughts as they come, Claude splits them into threads, and the project remembers how I work. It's where I do a ton of my coding now. [screenshot: my actual prompts for claude code cli yesterday] Quoting@ClaudeDevsToday we're rolling out Projects in Claude Code on desktop and web. A project is one conversation with Claude. It splits the work into threads itself, runs them as parallel cloud sessions, passes context between them, and keeps going when you leave. In beta for select users.
- François Chollet How Keras 3 helped modernise Expedia's ranking stack medium.com
- Boris Cherny Projects Recommendstheir own
- Sam Altman the main thing i was excited about launching this week will be next week instead, but imo worth the wait! Quoting@samabig 🚢 this week and then for devday 🚢🚢🚢🚢🚢🚢
- Boris Cherny Also today: Claude Docs, Claude Slides, and Claude Design are in every conversation. Ask Claude for a presentation and you get one you can open, edit, and export as PowerPoint or PDF. Same for a document or a design. There's no separate tool to navigate to. They're just in the chat. Quoting@claudeaiYou can also now make decks, docs, and designs in your conversation. Draft the one-pager in Claude Docs, turn it into a deck with Claude Slides, and mock up a matching visual in Claude Design, all from one place.
- Boris Cherny Claude Lovedtheir own
- Sam Altman big 🚢 this week and then for devday 🚢🚢🚢🚢🚢🚢 Quoting@thsottiauxThis week will also be a level of ships that you could have expected for DevDay 2025. Crazy
- François Chollet > One could even define intelligence as the efficiency with which one converts experience into competence; by this definition they lag very far behind us Yes, one could define intelligence in this way, and one would be right. By this metric current AI is approximately 6 OOMs less intelligent than humans. Your ancestors' evolutionary history did not prepare you for Python programming, yet you can learn to competently program in Python in a few hundreds of hours. An LRM needs the training data equivalent of ~1B hours (on top of all of its non-programming related training data). Another dimensio… Quoting@DKokotajloDan Selsam is a current OpenAI capabilities researcher. (since 2022) He was my boss for a while. He doesn't have a twitter account but has made this public statement of his views on AI risk and sent it to me to share: Dan Selsam's Personal Statement on AI Risk: I have been working on AI for over fi…
- Boris Cherny Claude Mods are landing now. Someone already built a Tetris-in-Claude mod 🤯 See issue for the latest community update, technical details, and more cool demos
- Sam Altman The world deserves confidence that American companies developing increasingly capable AI will act responsibly, especially as the trajectory of progress has steepened. Every frontier lab must deliver on this, and there is no reason any of us should come to work if we cannot. We welcome a federal framework that sets consistent safety requirements for frontier AI. But we do not believe we need to wait for an anti-trust exemption or legislation to begin the work of providing this confidence. Consistent rules to manage frontier risk so that we can maximize the benefits are a good idea (and we are…
- Sam Altman Alignment and safety techniques must stay ahead of progress in AI model capabilities.
- François Chollet Sincerity and self-interest usually go together by default.
- François Chollet People almost always align their beliefs with their interests.
- Boris Cherny Claude Liked
- François Chollet It is not rational to say AI is likely to end humanity in one's lifetime while completely winging safety measures.
- François Chollet If AI has a high chance of causing human extinction within years or decades, the only rational stance is stringent top-down government regulation and international treaties.
- François Chollet Current models are unsafe because they are not smart enough, not because they are too smart.
- François Chollet In the near term, more capable models should mean safer models.
- François Chollet The only way to avoid extreme power concentration is multiple independent providers of frontier AI models, including open-source options.
- François Chollet Extreme power concentration is one of the most worrying risks linked to frontier AI.
- François Chollet About a year ago, before it was on anyone's radar, we began exploring the idea of a benchmark for open-ended invention. Since then, we've developed several promising directions that will serve as the foundation for ARC 4 and ARC 5. We're incredibly excited to share what we've been building. We're still on track to release ARC 4 in Q1 next year, as promised. Quoting@arcprizeARC-AGI-4 will be a benchmark for autonomous open-ended innovation. It will continue our commitment to open-source, giving the research community a shared target for progress that benefits all of humanity. Despite rapid model progress, humans still significantly outperform AI at open-ended inventio…
- Andrej Karpathy I love this and really hope we can come together as an industry and make it happen. Quoting@DarioAmodeiWe Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they…
- Sam Altman I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon. Quoting@DarioAmodeiWe Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they…
- Boris Cherny Every day, I get a lot of of emails and messages like this one. I try to respond to as many as I can. Sharing my response below, for anyone else in a similar situation. What do you think?
- Boris Cherny Threat Intelligence report Recommends
- Sam Altman Welcome, Paul. Grateful you are doing this, and all you have done for AI safety. Excited to work together again.
- Sam Altman This would suck, but we will prioritize great service for customers until we can get back on top of things. Quoting@thsottiauxDemand for Astra is really unprecedented. We're pulling all the levers possible to sustain the demand, but I've not seen anything like it until now and we went through very steep growth before. Priority will always be to keep excellent service for existing users, but we might have to pause new Pro…
- Sam Altman Ocarina of Time Loved
- Sam Altman Mountain Dew Uses
- Sam Altman Ocarina of Time remake Loved
- Andrej Karpathy nanochat — The best ChatGPT that $100 can buy.
- Boris Cherny json-schema-to-typescript — Compile JSON Schema to TypeScript type declarations
- François Chollet GPT-6 Astra Loved
- Sam Altman An iterative loop where society and AI evolve together offers the best chance of getting AI development right.
- Sam Altman Managing the transition to a world with abundant and powerful AI should be one of the highest priorities in the world.
- Sam Altman No one fully understands the consequences of how capable AI is becoming.
- Boris Cherny Fable 5.1 Lovedtheir own
- Sam Altman This is a critically important moment for AI cyber defence and there is not much time left to act; only a collective response will work.
- François Chollet Deep Learning with Python Recommendstheir own
- Sam Altman Confidence in safety, not capability, will increasingly set the pace of AI progress.
- Sam Altman Keeping powerful models available only to a chosen few is not a good strategy.
-
François Chollet
The LLM line of research will reach a capability plateau.
On changing their mindIn the era of base LLM scaling (2022-2024), I believed the LLM line of research would reach a capability plateau (as later seen with base LLMs). In late 2024, after the o3 test-time compute demo, I changed my views: the new models were showing genuine fluid intelligence, and with this new line of work, the LLM line of research could achieve unbounded capability scaling. "There will be no wall."
- Boris Cherny Auto mode Usestheir own
- Sam Altman It is better to be an optimist who works hard than a pessimist who writes about why things will not work, because no amount of such essays moves society forward.
- Andrej Karpathy micrograd — A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API
- Andrej Karpathy We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for three js render of it. Opus went off for ~2 hours and wrote 5500 lines of code that (procedurally) rendered the story. It's kind of janky but fun. But it's a bit mindboggling that the LLM has to place and orchestrate various polygon assets in (x,y,z) coordinates and write code that animates it all, and that it even does…
- Boris Cherny A model is better thought of as a living creature with a personality you must get to know than as a component you specify.
- Boris Cherny Building on a model is unlike any previous software engineering, because the thing you are building on cannot be designed up front.
- Boris Cherny You cannot predict which instructions a model needs, so a system prompt has to be discovered by running the thing rather than designed in advance.
- Boris Cherny Every six months you should delete your whole AI configuration — instructions, skills and hooks — and find out how much of it the model never needed.
- Boris Cherny A frontier model is measurably more intelligent with no system prompt at all; the prompts that remain are there for the product, not the model.
- Boris Cherny Prompt work does not transfer across model generations: what you wrote three months ago for one model may be worth nothing to the next.
- Boris Cherny Prompt injection can be caught by watching the neurons that fire when it happens, so a defence no longer depends on the model reporting the attack.
- Boris Cherny Prompt injection is no longer a live attack on the frontier model: it simply does not follow instructions it reads on the internet any more.
- Boris Cherny Claude Opus 5 Likedtheir own
- Andrej Karpathy One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10 minutes, total mess, anything goes, full stream of consciousness. Sometimes I declare it up top, something like "switching to speech recognition sorry for any typos...". Sometimes I turn it into a small interview of a few turns. But I find that the LLMs are somehow very good at reconstructing long incoherent rambles and ofte…
- Boris Cherny Claude Code /checkup Recommendstheir own
- Andrej Karpathy This is a new paradigm for interacting with Claude that is significantly more "inline" with all the other human activity org-wide. Once you do all of the under the hood engineering work to make this "just work" (e.g. across tools, integrations, compute environments, memory, security, etc.), Claude basically joins the team in a seamless way - you can talk to it as you would talk to a person and it can help with a very large variety of workloads. Imo this is the 3rd major redesign of LLM UIUX. The first paradigm was that the LLM is a website you go to, the second was that it is an app you downl… Quoting@claudeaiIntroducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while you focus on other work.
- Andrej Karpathy In awe of SpaceX and its story - past, present and the future. You can think about it in 10+ different ways and continue re-blowing your mind in circles. Huge congrats to the team! 🚀
- Andrej Karpathy Claude Fable 5 Loved
- Andrej Karpathy Personal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team here and get back to R&D. I remain deeply passionate about education and plan to resume my work on it in time.
- Andrej Karpathy This works really well btw, at the end of your query ask your LLM to "structure your response as HTML", then view the generated file in your browser. I've also had some success asking the LLM to present its output as slideshows, etc. More generally, imo audio is the human-preferred input to AIs but vision (images/animations/video) is the preferred output from them. Around a ~third of our brains are a massively parallel processor dedicated to vision, it is the 10-lane superhighway of information into brain. As AI improves, I think we'll see a progression that takes advantage: 1) raw text (hard…
- Andrej Karpathy Sequoia Ascent 2026 summary Summary of my talk at Sequoia Ascent
- Andrej Karpathy karpathy.github.io — my blog
- Sam Altman - Here is a photo of my family. I love them more than anything. Images have power, I hope. Normally we try to be pretty private, but in this case I am sharing a photo in the hopes that it might dissuade the next person fr…
- Andrej Karpathy People disagreeing about AI capability are speaking past each other, because they are using models of very different tiers on very different kinds of task.
- Andrej Karpathy OpenAI Codex Liked
- Andrej Karpathy Claude Code Liked
- Andrej Karpathy AI will let citizens make their governments legible and accountable, reversing the historical direction in which only states could read society.
- Andrej Karpathy Because an LLM can argue almost any direction competently, the right way to use one for forming an opinion is to make it argue every side.
- Andrej Karpathy autoresearch — AI agents running research on single-GPU nanochat training automatically
- Andrej Karpathy LLM memory as currently built is a distraction to the model: one old question keeps resurfacing as if it were a lasting interest.
- Andrej Karpathy The apps built to drive smart devices should not exist at all: the devices should expose APIs and agents should call them directly.
- Andrej Karpathy The engineer's resource constraint has moved from flops to tokens: what matters now is the token throughput you command.
- Andrej Karpathy What limits you with coding agents is your own skill at stringing them together, not the capability of the models.
- Andrej Karpathy The way software gets built flipped in December: writing code yourself is now the exception, not the default.
- Andrej Karpathy jobs — A research tool for visually exploring Bureau of Labor Statistics Occupational Outlook Handbook data. This is not a report, a paper, or a serious economic publication — it is a development too…
- Andrej Karpathy microgpt This is a brief guide to my new art project microgpt, a single file of 200 lines of pure Python with no dependencies that trains and inferences a GPT. This file contains the full algorithmic content of what is needed: d…
- Andrej Karpathy rustbpe — The missing tiktoken training code
- Andrej Karpathy Nano Banana Loved
- Andrej Karpathy Poison Like No Other Recommends
- Andrej Karpathy In Defense of Food Recommends
- Andrej Karpathy Poison Squad Recommends
- Andrej Karpathy Auto-grading decade-old Hacker News discussions with hindsight A vibe coding thought exercise on what it might look like for LLMs to scour human historical data at scale and in retrospect.
- Andrej Karpathy hn-time-capsule — Analyzing Hacker News discussions from a decade ago in hindsight with LLMs
- Andrej Karpathy The space of minds On the space of minds and the optimizations that give rise to them.
- Andrej Karpathy llm-council — LLM Council works together to answer your hardest questions
- Andrej Karpathy Verifiability The impact of verifiability on the jagged frontier of LLMs
- Sam Altman Sora update #1 We have been learning quickly from how people are using Sora and taking feedback from users, rightsholders, and other interested groups. We of course spent a lot of time discussing this before launch, but now that we ha…
- Andrej Karpathy Animals vs Ghosts Today's frontier LLM research is not about building animals. It is about summoning ghosts. And a bit more on Sutton's Dwarkesh pod.
- Sam Altman Sora 2 We are launching a new app called Sora. This is a combination of a new model called Sora 2, and a new product that makes it easy to create, share, and view videos. This feels to many of us like the “ChatGPT for creativi…
- Sam Altman Abundant Intelligence Growth in the use of AI services has been astonishing; we expect it to be even more astonishing going forward. As AI gets smarter, access to AI will be a fundamental driver of the economy, and maybe eventually something…
- François Chollet deep-learning-with-python-notebooks — Jupyter notebooks for the code samples of the book "Deep Learning with Python"
- Sam Altman Jakub and Szymon AI has gotten remarkably better in recent years; ChatGPT can do amazing things that we take for granted. This is as it should be, and is the story of human progress. But behind the blinking circle, nicely abstracted awa…
- Boris Cherny redirect-claude
- Sam Altman The Gentle Singularity We are past the event horizon; the takeoff has started. Humanity is close to building digital superintelligence, and at least so far it’s much less weird than it seems like it should be. Robots are not yet walking the s…
- Boris Cherny anthropic-stream-repro
- François Chollet namex — Clean up the public namespace of your package!
- Boris Cherny undux — ⚡️ Dead simple state for React. Now with Hooks support.
- François Chollet ARC-AGI — The Abstraction and Reasoning Corpus
- Sam Altman Three Observations Our mission is to ensure that AGI (Artificial General Intelligence) benefits all of humanity. Systems that start to point to AGI* are coming into view, and so we think it’s important to understand the moment we are in.…
- Sam Altman Reflections The second birthday of ChatGPT was only a little over a month ago, and now we have transitioned into the next paradigm of models that can do complex reasoning. New years get people in a reflective mood, and I wanted to…
-
François Chollet
The memorize-fetch-apply paradigm behind LLMs can reach arbitrary skill given training data, but it cannot adapt to novelty or acquire new skills on the fly.
WasThis "memorize, fetch, apply" paradigm can achieve arbitrary levels of skills at arbitrary tasks given appropriate training data, but it cannot adapt to novelty or pick up new skills on the fly (which is to say that there is no fluid intelligence at play here.)
NowOpenAI's new o3 model represents a significant leap forward in AI's ability to adapt to novel tasks. This is not merely incremental improvement, but a genuine breakthrough, marking a qualitative shift in AI capabilities compared to the prior limitations of LLMs.
- Boris Cherny mcp-ping
- Boris Cherny es-module-stats — Collecting statistics on how the internet uses ES modules
- Boris Cherny json-schema-to-typescript-browser — Browser demo for json-schema-to-typescript
- François Chollet keras-resources — Directory of tutorials and open-source code repositories for working with Keras, the Python deep learning library
- François Chollet keras-blog — Blog with Keras news, tutorials, and demos.
- François Chollet nelder-mead — Pure Python/Numpy implementation of the Nelder-Mead algorithm.
- François Chollet deep-learning-models — Keras code and weights files for popular deep learning models.
- Sam Altman An Opinionated Guide to ML Research Recommends
- Sam Altman Emfit QS+Active Liked
- Sam Altman Chili Pad Mixed on
- Sam Altman full spectrum LED light Recommends
- François Chollet hualos — Keras Total Visualization project