Episode notes
Eric Jang walks through how to build AlphaGo from scratch, but with modern AI tools. Sometimes you understand the future better by stepping backward. AlphaGo is still the cleanest worked example of the primitives of intelligence: search, learning from experience, and self-play. You have to go back to 2017 to get insight into how the more general AIs of the future might learn. Once he explained how AlphaGo works, it gave us the context to have a discussion about how RL works in LLMs and how it could work better – naive policy gradient RL has to figure out which of the 100k+ tokens in your…
Dwarkesh Podcast
by Dwarkesh Patel · English · Tech & Science
Deeply researched interviews www.dwarkesh.com
More from Dwarkesh Podcast
-
16 Jun 2026 · 2 hr 8 min
Ada Palmer – Machiavelli is the most misunderstood thinker of all time
Had Ada Palmer back on – this time to talk about Machiavelli, perhaps the most misunderstood thinker of all time. Machiavelli cut his teeth as a high-level diplomat for Florence, a position from which he got to closely observe the most important rulers in Europe at the time, including the ones who were on the path to destroying his dearly beloved Florence. In 1513 the Medici retook control of Florence and, wrongly suspecting Machiavelli of participating in a coup attempt, fired, tortured, and exiled him. Machiavelli could have left exile and worked for any number of different principalities…
-
4 Jun 2026 · 1 hr 16 min
Alex Imas and Phil Trammell – What remains scarce after AGI?
Economics of AGI episode w Alex Imas and Phil Trammell . There’s a bunch of important questions about how we deal with AI that only economics can answer. What is the optimal way to tax and redistribute the wealth that will be generated? How should countries not in the AI supply chain index into the gains? Is there any world where inequality doesn’t explode? It might seem like these questions have obvious answers, but the first thing economics teaches you is that your intuitions can often be entirely wrong. It was very helpful to chat through these things with Alex and Phil. Watch on YouTube…
-
22 May 2026 · 1 hr 21 min
Reiner Pope – Chip design from the bottom up
New blackboard lecture with Reiner Pope: how do chips actually work - starting with basic logic gates, and working up to why GPUs, TPUs, FPGAs, and the human brain each look the way they do. Reiner is CEO of MatX , a new chip startup (full disclosure - I’m an angel investor). He was previously at Google, where he worked on software efficiency , compilers, and TPU architecture. Watch this one on YouTube so you can see the chalkboard. Read the transcript . Sponsors * Crusoe was one of only five GPU clouds that made the gold tier in SemiAnalysis' most recent ClusterMAX report. Gold-tier…
-
8 May 2026 · 2 hr 13 min
David Reich – Why the Bronze Age was an inflection point in human evolution
David Reich is back. He and collaborator Ali Akbari just published a paper that overturns a long-standing consensus about human evolution — that natural selection has been dormant in our species since the agricultural revolution. By scaling ancient DNA sequencing and developing a new statistical method, they found that selection has actually sped up. Selection went especially bonkers during the Bronze Age (around 3,000 years ago). That’s when gene frequencies for everything from immune function to body fat to intelligence were most in flux. Over the last 10,000 years, selection pushed the…
-
29 Apr 2026 · 2 hr 14 min
Reiner Pope – The math behind how LLMs are trained and served
Did a very different format with Reiner Pope - a blackboard lecture where he walks through how frontier LLMs are trained and served. It’s shocking how much you can deduce about what the labs are doing from a handful of equations, public API prices, and some chalk. It’s a bit technical, but I encourage you to hang in there – it’s really worth it. There are less than a handful of people who understand the full stack of AI, from chip design to model architecture, as well as Reiner. It was a real delight to learn from him. Recommend watching this one on YouTube so you can see the chalkboard.…
-
15 Apr 2026 · 1 hr 43 min
Jensen Huang – TPU competition, why we should sell chips to China, & Nvidia’s supply chain moat
I asked Jensen about TPU competition, Nvidia’s lock on the ever more bottlenecked supply chain needed to make advanced chips, whether we should be selling AI chips to China, why Nvidia doesn’t just become a hyperscaler, how it makes its investments, and much more. Enjoy! Watch on YouTube ; read the transcript . Sponsors * Crusoe’s cloud runs on state-of-the-art Blackwell GPUs, with Vera Rubin deployment scheduled for later this year. But hardware is only part of the story—for inference, Crusoe’s MemoryAlloy tech implements a cluster-wide KV cache, delivering up to 10x faster TTFT and 5x…
-
1 Oct 2026 · 1 hr 37 min
Si Sheppard – How did a few hundred Spanish soldiers topple two empires?
New episode with military historian Si Sheppard . The conquests of the Aztec and Inca empires might be the most shocking events in human history. Hernan Cortés and some hundreds of Conquistadors landed on a new continent they knew basically nothing about, and within 2 and a half years, conquered the 6 million strong Aztec empire. Then, a decade later, Francisco Pizarro did the same thing to an Incan Empire of some 10 million inhabitants. In both cases, the Conquistadors consistently beat native armies while being literally 100x (or even in some cases 1000x) outnumbered - and in many cases,…
-
17 Sep 2026 · 1 hr 20 min
Noam Brown – Agent swarms, alignment, & recursive self-improvement
New episode with Noam Brown. We talk about multi-agent, Navier-Stokes, and what the current explosion of maths progress tells us about what happens once you automate AI research. And we also discuss how we will know if the models are actually aligned before we kick off RSI. Watch on YouTube ; read the transcript . Sponsors * Jane Street has been interested in AI for a lot longer than you’d think, and not just for trading. In 2011, a full year before AlexNet and over a decade before ChatGPT launched, they hosted the first FOOM Debate between Eliezer Yudkowsky and Robin Hanson on whether AI…
-
11 Sep 2026 · 1 hr 37 min
AI researchers debate how close we are to recursive self-improvement
New episode with John Schulman , Beren Millidge and Charlie O’Neill . I got together with some of the most insightful AI researchers I know who are at the openish companies, because I wanted to hear the details of what's actually happening at the frontier and what comes next. Watch on YouTube ; read the transcript . Sponsors * Antithesis helps you trust your code. As agents generate more and more of your software, the bottleneck shifts from your engineers actually writing code to verifying it. Antithesis does that testing for you. Ron Minsky, who co-leads Jane Street’s tech group, told me…
-
1 Sep 2026 · 2 hr 21 min
Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Face
Ajeya Cotra is a researcher at METR, where she works on threat modeling for loss-of-control risks from advanced AI. Before that, she led the technical AI safety program at what is now Coefficient Giving. She is one the three authors of METR and Redwood Research’s “ Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident ”. We go through not only what she and her coauthors discovered during this investigation, but what it means for how we should train future, smarter AIs which might be involved in the process of recursive…
