{"podcast":{"title":"The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)","slug":"twiml-ai-podcast","podcast_index_feed_id":1045879,"rss_url":"https://feeds.megaphone.fm/MLN2155636147","website_url":"https://twimlai.com","image_url":"https://megaphone.imgix.net/podcasts/35230150-ee98-11eb-ad1a-b38cbabcd053/image/TWIML_AI_Podcast_Official_Cover_Art_1400px.png?ixlib=rails-4.3.1&max-w=3000&max-h=3000&fit=crop&auto=format,compress","author":"TWIML","episode_count":785,"summary":"Machine learning and artificial intelligence are dramatically changing the way businesses operate and people live. The TWIML AI Podcast brings the top minds and ideas from the world of ML and AI to a broad and influential community of ML/AI researchers, data scientists, engineers and tech-savvy business and IT leaders. Hosted by Sam Charrington, a sought after industry analyst, speaker, commentator and thought leader. Technologies covered include machine learning, artificial intelligence, deep learning, natural language processing, neural networks, analytics, computer science, data science and more.","last_synced_at":null,"page_url":"https://stenobird.com/podcast/twiml-ai-podcast"},"episode":{"title":"Infrastructure Scaling and Compound AI Systems with Jared Quincy Davis - #740","slug":"infrastructure-scaling-and-compound-ai-systems-with-jared-quincy-davis-740","published_at":"2025-07-22T16:00:00+00:00","page_url":"https://stenobird.com/podcast/twiml-ai-podcast/infrastructure-scaling-and-compound-ai-systems-with-jared-quincy-davis-740","show_page_url":"https://stenobird.com/podcast/twiml-ai-podcast","url":"https://twimlai.com/podcast/twimlai/infrastructure-scaling-and-compound-ai-systems/","audio_url":"https://pscrb.fm/rss/p/traffic.megaphone.fm/MLN2657285858.mp3?updated=1753151737","summary":"In this episode, Jared Quincy Davis, founder and CEO at Foundry, introduces the concept of \"compound AI systems,\" which allows users to create powerful, efficient applications by composing multiple, often diverse, AI models and services. We discuss how these \"networks of networks\" can push the Pareto frontier, delivering results that are simultaneously faster, more accurate, and even cheaper than single-model approaches. Using examples like \"laconic decoding,\" Jared explains the practical techniques for building these systems and the underlying principles of inference-time scaling. The conversation also delves into the critical role of co-design, where the evolution of AI algorithms and the underlying cloud infrastructure are deeply intertwined, shaping the future of agentic AI and the compute landscape. The complete show notes for this episode can be found at https://twimlai.com/go/740.","meta_description":"In this episode, Jared Quincy Davis, founder and CEO at Foundry, introduces the concept of \"compound AI systems,\" which allows users to create powerful, e…","key_points":[],"chapters":[],"topics":[],"duration_seconds":4382,"processing_state":"failed","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/twiml-ai-podcast/episodes/infrastructure-scaling-and-compound-ai-systems-with-jared-quincy-davis-740/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/twiml-ai-podcast/infrastructure-scaling-and-compound-ai-systems-with-jared-quincy-davis-740.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}