Steadcast
Dwarkesh Podcast cover art
Dwarkesh Podcast

Reiner Pope – The math behind how LLMs are trained and served

April 29, 20262h 13m · 22,975 words

Show notes

Did a very different format with Reiner Pope - a blackboard lecture where he walks through how frontier LLMs are trained and served. It’s shocking how much you can deduce about what the labs are doing from a handful of equations, public API prices, and some chalk. It’s a bit technical, but I encourage you to hang in there – it’s really worth it.

Highlighted moments

if you do not batch together many users, the cost and the economics you get can be like a thousand times worse than if you do batch many two users together.
The fact that they are charging 5x less for pre-fill than decode does suggest that they are bottlenecked on memory bandwidth to quite a degree

Transcript

Transcript not available for this episode yet.

More from Dwarkesh Podcast

8 Predictions for the Era of Continual Learning

Aug 7, 20268 min

Why smarter AI models could drive up compute prices 10x

Aug 3, 202611 min

Adam Brown – A deep but accessible introduction to general relativity

Jul 10, 20261h 38m

Grant Sanderson – AI and the future of math

Jun 30, 20261h 33m

The next big breakthrough will be AIs learning on the job

Jun 26, 202619 min