Episode

Product Metrics are LLM Evals // Raza Habib CEO of Humanloop // #320

Podcast
MLOps.community
Published
Jun 3, 2025
Duration seconds
3186
Processing state
processed
Canonical source
https://podcasters.spotify.com/pod/show/mlops/episodes/Product-Metrics-are-LLM-Evals--Raza-Habib-CEO-of-Humanloop--320-e33moi9
Audio
https://anchor.fm/s/174cb1b8/podcast/play/103555081/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2025-5-2%2F401451062-44100-2-a2eaaf386212f.mp3
JSON
/v1/public/podcasts/mlops-community/episodes/product-metrics-are-llm-evals-raza-habib-ceo-of-humanloop-320
Markdown
/podcast/mlops-community/product-metrics-are-llm-evals-raza-habib-ceo-of-humanloop-320.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/mlops-community/episodes/product-metrics-are-llm-evals-raza-habib-ceo-of-humanloop-320/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/mlops-community/product-metrics-are-llm-evals-raza-habib-ceo-of-humanloop-320.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Raza Habib, the CEO of the LLM Eval platform Humanloop , talks to us about how to make your AI products more accurate and reliable by shortening the feedback loop of your evals. Quickly iterating on prompts and testing what works, along with some of his favorite Dario from Anthropic AI Quotes. // Bio Raza is the CEO and Co-founder at Humanloop. He has a PhD in Machine Learning from UCL, was the founding engineer of Monolith AI, and has built speech systems at Google. For the last 4 years, he has led Humanloop and supported leading technology companies such as Duolingo, Vanta, and Gusto to build products with large language models. Raza was featured in the Forbes 30 Under 30 technology list in 2022, and Sifted recently named him one of the most influential Gen AI founders in Europe. // Related Links Websites: https://humanloop.com ~~~~~~~~ ✌️Connect With Us ✌️ ~~~~~~~ Catch all episodes, blogs, newsletters, and more: https://go.mlops.community/TYExplore MLOps Swag/Merch: [ https://shop.mlops.community/ ] Connect with Demetrios on LinkedIn: /dpbrinkm Connect with Raza on LinkedIn: /humanloop-raza Timestamps: [00:00] Cracking Open System Failures and How We Fix Them [05:44] LLMs in the Wild — First Steps and Growing Pains [08:28] Building the Backbone of Tracing and Observability [13:02] Tuning the Dials for Peak Model Performance [13:51] From Growing Pains to Glowing Gains in AI Systems [17:26] Where Prompts Meet Psychology and Code [22:40] Why Data Experts Deserve a Seat at the Table [24:59] Humanloop and the Art of Configuration Taming [28:23] What Actually Matters in Customer-Facing AI [33:43] Starting Fresh with Private Models That Deliver [34:58] How LLM Agents Are Changing the Way We Talk [39:23] The Secret Lives of Prompts Inside Frameworks [42:58] Streaming Showd…