
technologyJun 30, 20261:10:14pending
Why Hardware-Software Co-Design Is AI's Real 100x: Dylan Patel of SemiAnalysis
About this episode
Dylan Patel, founder of SemiAnalysis, argues the biggest gains in AI don't come from faster chips, they come from software-hardware co-design. Optimizing the model, the kernels, and the silicon together turns a 2x here and a 2x there into 100x. He explains why DeepSeek's experts were shaped for Nvidia's Hopper (and why TPUs struggle to run it), why OpenAI's sparser models and Anthropic's denser ones pull them toward different hardware, and why the so-called CUDA moat was never really about CUDA. Dylan breaks down InferenceX, his living benchmark that runs the latest models on over $50M of donated hardware daily, tracking a roughly 60x annual drop in cost per unit of quality. He makes the case that inference will be a bigger market than oil, that the compute crunch persists because models expand the value of useful work faster than compute grows, and why Jensen Huang is bankrolling neoclouds to engineer a multipolar world.
Hosted by Shaun Maguire and Sonya Huang, Sequoia Capital
Get every episode summarized
Each time Training Data publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Training Data

Making Cities Awesome: Peregrine’s Nick Noone & Ben Rudolph
Training Data
Sep 1, 202652:27pending

Parallel’s Parag Agrawal: Building a New Web for AI Agents
Training Data
Aug 25, 202655:18pending

Rich Sutton and Khurram Javed: Why AI Models Stop Learning, and How to Start It...
Training Data
Aug 18, 202653:43pending

Chai Discovery's Bitter Lesson: Drug Design Is Another Scaling Problem
Training Data
Aug 4, 202647:22pending