Andrej Karpathy
Everything, newest first — across every channel. Their profile →
Filter & sortAll sources · condensed
- micrograd — A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API 3 Aug
- nanochat — The best ChatGPT that $100 can buy. 2 Aug
- We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested w… 2 Aug
- One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, b… 21 Jul
- This is a new paradigm for interacting with Claude that is significantly more "inline" with all the other human activity org-wide. Once you do all of the under… 23 Jun
- In awe of SpaceX and its story - past, present and the future. You can think about it in 10+ different ways and continue re-blowing your mind in circles. Huge… 12 Jun
- This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great and it's SOTA on e… 9 Jun
- Personal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team her… 19 May
- This works really well btw, at the end of your query ask your LLM to "structure your response as HTML", then view the generated file in your browser. I've also… 11 May
- This is the the quote I've been citing a lot recently. 30 Apr
- Fireside chat at Sequoia Ascent 2026 from a ~week ago. Some highlights: The first theme I tried to push on is that LLMs are about a lot more than just speeding… 30 Apr
- Sequoia Ascent 2026 summary 30 Apr
- karpathy.github.io — my blog 10 Apr
- Judging by my tl there is a growing gap in understanding of AI capability. The first issue I think is around recency and tier of use. I think a lot of people t… 9 Apr
- Farzapedia, personal wikipedia of Farza, good example following my Wiki LLM tweet. I really like this approach to personalization in a number of ways, compared… 4 Apr
- Something I've been thinking about - I am bullish on people (empowered by AI) increasing the visibility, legibility and accountability of their governments. Hi… 4 Apr
- Wow, this tweet went very viral! I wanted share a possibly slightly improved version of the tweet in an "idea file". The idea of the idea file is that in this… 4 Apr
- LLM Knowledge Bases Something I'm finding very useful recently: using LLMs to build personal knowledge bases for various topics of research interest. In this w… 2 Apr
- New supply chain attack this time for npm axios, the most popular HTTP client library with 300M weekly downloads. Scanning my system I found a use imported fro… 31 Mar
- - Drafted a blog post - Used an LLM to meticulously improve the argument over 4 hours. - Wow, feeling great, it’s so convincing! - Fun idea let’s ask it to arg… 28 Mar
- When I built menugen ~1 year ago, I observed that the hardest part by far was not the code itself, it was the plethora of services you have to assemble like IK… 26 Mar
- autoresearch — AI agents running research on single-GPU nanochat training automatically 26 Mar
- One common issue with personalization in all LLMs is how distracting memory seems to be for the models. A single question from 2 months ago about some topic ca… 25 Mar
- Software horror: litellm PyPI supply chain attack. Simple `pip install litellm` was enough to exfiltrate SSH keys, AWS/GCP/Azure creds, Kubernetes configs, git… 24 Mar
- Thank you Sarah, my pleasure to come on the pod! And happy to do some more Q&A in the replies. 21 Mar
- jobs — A research tool for visually exploring Bureau of Labor Statistics Occupational Outlook Handbook data. This is not a report, a paper, or a serious economic publication — it is a development too… 16 Mar
- Outages at the frontier AI labs will become "intelligence brownouts" — the planet losing IQ points when the models stutter. 11 Mar
- Agentic organisations will be forkable in a way that classical companies never were. 11 Mar
- The age of the IDE is not over; agents make it bigger, because humans now program at a higher level where the unit of interest is an agent rather than a file. 11 Mar
- microgpt 12 Feb
- rustbpe — The missing tiktoken training code 3 Jan
- Chemical hygiene 22 Dec 2025
- 2025 LLM Year in Review 20 Dec 2025
- Auto-grading decade-old Hacker News discussions with hindsight 10 Dec 2025
- hn-time-capsule — Analyzing Hacker News discussions from a decade ago in hindsight with LLMs 10 Dec 2025
- The space of minds 29 Nov 2025
- llm-council — LLM Council works together to answer your hardest questions 22 Nov 2025
- reader3 — Quick illustration of how one can easily read books together with LLMs. It's great and I highly recommend it. 18 Nov 2025
- Verifiability 17 Nov 2025
- The best historical analogy for AI is not electricity or the industrial revolution but a new computing paradigm, because both are fundamentally about automating digital information processing. 16 Nov 2025
- nanoGPT — The simplest, fastest repository for training/finetuning medium-sized GPTs. 12 Nov 2025
- Animals vs Ghosts 1 Oct 2025
- Most current predictions about AI's imminent impact on the job market are naive. 25 Sept 2025
- Predictions that AI would eliminate radiology jobs were wrong; radiology is growing. 25 Sept 2025
- rendergit — Render any git repo into a single static HTML page for humans or LLMs 21 Aug 2025
- llm.c — LLM training in simple, raw C/CUDA 26 Jun 2025
- Vibe coding MenuGen 1 May 2025
- Power to the people: How LLMs flip the script on technology diffusion 8 Apr 2025
- Finding the Best Sleep Tracker 25 Mar 2025
- nn-zero-to-hero — Neural Networks: Zero to Hero 18 Aug 2024
- minGPT — A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training 15 Aug 2024
- build-nanogpt — Video+code lecture on building nanoGPT from scratch 13 Aug 2024
- llama2.c — Inference Llama 2 in one file of pure C 6 Aug 2024
- LLM101n — LLM101n: Let's build a Storyteller 1 Aug 2024
- minbpe — Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization. 1 Jul 2024
- makemore — An autoregressive character-level language model for making more things 4 Jun 2024
- Deep Neural Nets: 33 years ago and 33 years from now 14 Mar 2022
- A from-scratch tour of Bitcoin in Python 21 Jun 2021
- Short Story on AI: Forward Pass 27 Mar 2021
- Biohacking Lite 11 Jun 2020
- A Recipe for Training Neural Networks 25 Apr 2019
- (started posting on Medium instead) 20 Jan 2018
- A Survival Guide to a PhD 7 Sept 2016
- Deep Reinforcement Learning: Pong from Pixels 31 May 2016
- Short Story on AI: A Cognitive Discontinuity. 14 Nov 2015