

Large models on CPUs
Model sizes are crazy these days with billions and billions of parameters. As Mark Kurtz explains in this episode, this makes inference slow and expensive despite the fact that up to 90%+ of the parameters don’t influence the outputs at all. Mark helps us understand all of the practicalities and progress that is being…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
Similar episodes from other shows
More like Practical AI →
LLMs Everywhere: Running 70B models in browsers and iPhones using MLC — with Tianqi Chen of CMU / OctoML

What People Are Actually Using AI For Right Now

Reiner Pope – The math behind how LLMs are trained and served

#324 Sharon Zhou: Inside AMD's Plan to Build Self-Improving AI
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.


