Podcast cover
Hamel Husain · Technology

Hamel Husain

I am a machine learning engineer with over 20 years of experience. I have worked with innovative companies such as Airbnb and GitHub, which included early LLM research used by OpenAI, for code understanding. I have also led and contributed to numerous popular open-source machine-learning tools. I am currently an independent consultant helping companies operationalize Large Language Models (LLMs) to accelerate their AI product journey.

Episodes

Episode cover

How To Make PDFs Your AI Can Actually Read

02 Oct 2026
8m
Episode cover

Giving AI a Credit Card Is Still Painful

30 Sep 2026
8m
Episode cover

8 Claude Code skills I use for building better AI evals

30 Sep 2026
11m
Episode cover

Why The Prompt Matters Less Than The Context

28 Sep 2026
5m
Episode cover

How to Apply Data Science Skills to AI Engineering

25 Sep 2026
4m
Episode cover

How to Design a Data Agent People Can Verify

22 Sep 2026
6m
Episode cover

Putting Claudes "AI Slop" Solution to the Test

15 Sep 2026
7m
Episode cover

Trying the new Claude Eval tool

12 Sep 2026
58m
Episode cover

How To Build And Evaluate Search Agents

03 Sep 2026
50m
Episode cover

How to Stop Building Products Nobody Wants

02 Sep 2026
1h 31m
AI processed

Continuous discovery transforms product development by replacing speculative, idea-first roadmaps with structured, customer-centric feedback loops. Product management expert Teresa Torres emphasizes that effective discovery requires moving beyond superficial customer interviews—which often trigger confirmation bias—tow...

Episode cover

Stop Picking Embedding Models Off The MTEB Leaderboard

31 Aug 2026
21m
Episode cover

Don't Build Agents, Build Environments Instead

25 Aug 2026
27m
Episode cover

How Multi-Vector Retrieval Works at Scale

21 Aug 2026
24m
Episode cover

How To Turn Evals Into A Better Model

17 Aug 2026
35m
Episode cover

How to Cut Your LLM Classification Costs by 90%

12 Aug 2026
25m
Episode cover

How To Use Open Models Effectively

07 Aug 2026
40m
Episode cover

How to Build Agents That Answer Data Questions

31 Jul 2026
27m
Episode cover

How To Make Codex Run Itself

27 Jul 2026
4m
Episode cover

How To Choose The Right OCR Model

24 Jul 2026
23m
Episode cover

How To Build AI Evals

17 Jul 2026AI processed

Implementing robust evaluation systems for AI agents requires a systematic approach that prioritizes observability and high-quality, human-labeled data. Lucas, a developer at Nova Escola, details his transition from manual, spreadsheet-based assessments to an automated pipeline integrated with Claude and Langfuse. By i...

Page 1

Follow this podcast in Podwise

Sign in to get AI summaries, transcripts and mind maps for any episode, including new ones.

Open in Podwise