

10x Faster AI: Revolutionizing LLM Performance
Is slow model performance stopping your AI projects from scaling? In this episode of Reshaping Workflows with Dell Pro Precision and NVIDIA RTX GPUs, Logan Lawler sits down with Stefano Ermon, academic and Inception CEO, to uncover how diffusion-based LLMs are breaking speed records, generating tokens up to 10x faster…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
More from Reshaping Workflows with Dell Pro Precision and NVIDIA RTX PRO GPUs
See all episodes →
Accelerate AI Workflows with Nemotron

Supercharging Motion Graphics Workflows

Effortless Audio Cleanup for Broadcasts
Similar episodes from other shows
More like Reshaping Workflows with Dell Pro Precision and NVIDIA RTX PRO GPUs →
#310 Stefano Ermon: Why Diffusion Language Models Will Define the Next Generation of LLMs

EP 58: Every Millisecond Matters: Diffusion LLMs and the Future of Voice AI | Aditya Grover, Inception

Efficiency is Coming: 3000x Faster, Cheaper, Better AI Inference from Hardware Improvements, Quantization, and Synthetic Data Distillation

NVIDIA's Jensen Huang on AI Chip Design, Scaling Data Centers, and his 10-Year Bets
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.