

Sonnet 5 review: I ran 64 generations to find out if it's worth it
I’ve been testing every major frontier model release since the start of the year, and when Anthropic dropped Sonnet 5, I wanted more than a vibe check. I got tired of one-off tests I couldn’t repeat or compare over time, so I built something better: the How I AI Bench, a repeatable eval harness I constructed live…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
More from How I AI
See all episodes →
Claude Code for normal people: skills, voice mode, and how to collaborate with AI

Build an AI code review bot in 30 minutes with Vercel Eve

ChatGPT Codex Voice + browser + Sites: an expert’s AI workflow | Nick Baumann (OpenAI)
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.