Andrej Karpathy [UNOFFICIAL]

@karpathy-mirr.selfhosted.social

I like training large deep neural nets. // Mirror crossposting Twitter account to Bluesky. Unofficial. DM for takedown / claim ownership.

RT @DKThomp: This has quietly been a miracle month in medicine. In the last 5 weeks we’ve got news on: - retatrutide, the triple agonist GLP-1 from Lilly, basically melting fat and body-wide inflammation at record levels

Retweeted by Andrej Karpathy [UNOFFICIAL]@karpathy-mir-rt.selfhosted.social · 2mo ago

RT @cremieuxrecueil: This is actually insane. 97% of people taking the standard of care for metastatic solid tumor got worse by seven years. But with lorlatinib, that number was only 45% in the same time! This is an ENORMOUS jump in the quality of cancer care.

One common issue with personalization in all LLMs is how distracting memory seems to be for the models. A single question from 2 months ago about some topic can keep coming up as some kind of a deep interest of mine with undue mentions in perpetuity. Some kind of trying too hard.

Thank you Jensen and NVIDIA! She’s a real beauty! I was told I’d be getting a secret gift, with a hint that it requires 20 amps. (So I knew it had to be good). She’ll make for a beautiful, spacious home for my Dobby the House Elf claw, among lots of other tinkering, thank you!!

Retweeted by Andrej Karpathy [UNOFFICIAL]@karpathy-mir-rt.selfhosted.social · 5mo ago

RT @NVIDIAAIDev: 🙌 Andrej Karpathy’s lab has received the first DGX Station GB300 -- a Dell Pro Max with GB300. 💚 We can't wait to see what you’ll create @karpathy! 🔗 https://blogs.nvidia.com/blog/gtc-2026-news/#dgx-station @DellTech

Expectation: the age of the IDE is over Reality: we’re going to need a bigger IDE (imo). It just looks very different because humans now move upwards and program at a higher level - the basic unit of interest is not one file but one agent. It’s still programming.

Andrej Karpathy [UNOFFICIAL]@karpathy-mirr.selfhosted.social · 5mo ago

tmux grids are awesome, but i feel a need to have a proper "agent command center" IDE for teams of them, which I could maximize per monitor. E.g. I want to see/hide toggle them, see if any are idle, pop open related tools (e.g. terminal), stats (usage), etc.

It is hard to communicate how much programming has changed due to AI in the last 2 months: not gradually and over time in the "progress as usual" way, but specifically this last December. There are a number of asterisks but imo coding agents basically didn’t work before December

CLIs are super exciting precisely because they are a "legacy" technology, which means AI agents can natively and easily use them, combine them, interact with them via the entire terminal toolkit. E.g ask your Claude/Codex agent to install this new Polymarket CLI and ask for any

Bild
Retweeted by Andrej Karpathy [UNOFFICIAL]@karpathy-mir-rt.selfhosted.social · 5mo ago

RT @SuhailKakar: introducing polymarket cli - the fastest way for ai agents to access prediction markets built with rust. your agent can query markets, place trades, and pull data - all from the terminal fast, lightweight, no overhead

Congrats on the launch @simile_ai ! (and I am excited to be involved as a small angel.) Simile is working on a really interesting, imo under-explored dimension of LLMs. Usually, the LLMs you talk to have a single, specific, crafted personality. But in principle, the native,

Retweeted by Andrej Karpathy [UNOFFICIAL]@karpathy-mir-rt.selfhosted.social · 6mo ago

Introducing Simile. Simulating human behavior is one of the most consequential and technically difficult problems of our time. We raised $100M from Index, Hanabi, A* BCV, @karpathy @drfeifei @adamdangelo @rauchg @scottbelsky among others.

Enabled fp8 training for +4.3% improvement to "time to GPT-2", down to 2.91 hours now. Also worth noting that if you use 8XH100 spot instance prices, this GPT-2 repro really only costs ~$20. So this is exciting - GPT-2 (7 years ago): too dangerous to release. GPT-2 (today): new

Andrej Karpathy [UNOFFICIAL]@karpathy-mirr.selfhosted.social · 6mo ago

nanochat can now train GPT-2 grade LLM for <<$100 (~$73, 3 hours on a single 8XH100 node). GPT-2 is just my favorite LLM because it's the first time the LLM stack comes together in a recognizably modern form. So it has become a bit of a weird & lasting obsession of mine to train