

Sonnet 5 review: I ran 64 generations to find out if it's worth it
I’ve been testing every major frontier model release since the start of the year, and when Anthropic dropped Sonnet 5, I wanted more than a vibe check. I got tired of one-off tests I couldn’t repeat or compare over time, so I built something better: the How I AI Bench, a repeatable eval harness I constructed live…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
More from How I AI
See all episodes →Similar episodes from other shows
More like How I AI →
EP 301: Anthropic Claude 3.5 Sonnet – How it compares to ChatGPT's GPT-4o

Sonnet 4.6 Changes the Agent Math

AI Fundamentals: Benchmarks 101
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.


