Episode

Interview #86 Carter Huffman, CTO of Modulate

Podcast
The Artificial Intelligence Podcast
Published
May 7, 2026
Duration seconds
1840
Processing state
not_requested
Canonical source
https://podcasters.spotify.com/pod/show/tonyphoang/episodes/Interview-86-Carter-Huffman--CTO-of-Modulate-e3j0ic1
Audio
https://traffic.megaphone.fm/APO1551725216.mp3
JSON
/v1/public/podcasts/the-artificial-intelligence-podcast-3635848/episodes/interview-86-carter-huffman-cto-of-modulate
Markdown
/podcast/the-artificial-intelligence-podcast-3635848/interview-86-carter-huffman-cto-of-modulate.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-artificial-intelligence-podcast-3635848/episodes/interview-86-carter-huffman-cto-of-modulate/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848/interview-86-carter-huffman-cto-of-modulate.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Carter Huffman, CTO of Modulate, breaks down why voice AI keeps failing in production despite years of "solved" transcription claims. He explains the "transcript trap" that fools dev teams into shipping voice agents that crumble on real calls, why bigger language models actually make latency worse, and how ensemble approaches with smaller specialized models outperform monolithic systems. Carter also dives into the explosion of deepfake voice fraud - including the $25M Hong Kong heist - and why passive monitoring plus red teaming are now essential to a modern voice security stack. Finally, he shares why even Gen Z still defaults to voice for high-stakes issues, and what contact centers must get right over the next three years to build genuine trust with callers.