SkepSkep
All episodes
Latent Space: The AI Engineer Podcast — FlashAttention 2: making Transformers 800% faster w/o approximation - with Tri Dao of Together AI
Not scored yet54m·2023-07-26·Technology

FlashAttention 2: making Transformers 800% faster w/o approximation - with Tri Dao of Together AI

Latent Space: The AI Engineer Podcast

FlashAttention was first published by Tri Dao in May 2022 and it had a deep impact in the large language models space. Most open models you’ve heard of (RedPajama, MPT, LLaMA, Falcon, etc) all leverage it for faster inference. Tri came on the podcast to chat about FlashAttention, the newly released FlashAttention-2,…

Listen in your app
Buzzmeter · cross-platform
Not scored yet

We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.

Hive Vote · Skep listeners
New on SkepAwaiting the hive's first votes

Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.

Rate this episode

Add your vote to the Hive — listeners rate every episode after they finish.

Rate this episode
Recent reviews
  • No reviews yet — be the first.

One great episode in your inbox, daily — free.

One email a day, unsubscribe anytime. No spam, ever.