

Evals, error analysis, and better prompts: A systematic approach to improving your AI products | Hamel Husain (ML engineer)
Hamel Husain, an AI consultant and educator, shares his systematic approach to improving AI product quality through error analysis, evaluation frameworks, and prompt engineering. In this episode, he demonstrates how product teams can move beyond “vibe checking” their AI systems to implement data-driven quality…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
More from How I AI
See all episodes →
Claude Code for normal people: skills, voice mode, and how to collaborate with AI

Build an AI code review bot in 30 minutes with Vercel Eve

ChatGPT Codex Voice + browser + Sites: an expert’s AI workflow | Nick Baumann (OpenAI)
Similar episodes from other shows
More like How I AI →
Why AI evals are the hottest new skill for product builders | Hamel Husain & Shreya Shankar (creators of the #1 eval course)

Collaboration & evaluation for LLM apps

Building the Foundation Model Ops Platform — with Raza Habib of Humanloop

How to Prompt GPT-5
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.