# Reconfigurable Hardware: ElastixAI and The Future of Fast, Efficient AI Inference Page: https://stenobird.com/podcast/amelia-s-weekly-fish-fry-166630/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference Text version: https://stenobird.com/podcast/amelia-s-weekly-fish-fry-166630/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference.md Podcast: [Amelia's Weekly Fish Fry](https://stenobird.com/podcast/amelia-s-weekly-fish-fry-166630) Published: 2026-06-12T07:00:00+00:00 Episode link: https://techfocus.podbean.com/e/how-elastic-ai-uses-fpgas-to-crush-llm-inference-bottlenecks/ Audio file: https://mcdn.podbean.com/mf/web/3mfuzdmdir43qvby/Fish_fry_-_June_12_20269fr29.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/amelia-s-weekly-fish-fry-166630/episodes/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference Duration seconds: 1143 ## Resource Artificial intelligence is moving faster than ever, but as AI models continue to grow in size and complexity, the challenges surrounding inference performance are becoming impossible to ignore. In this week's podcast, ElastixAI CEO Dr. Mohammad Rastegari and I chat about how we can overcome those challenges and why a different approach to AI infrastructure is necessary for the next generation of AI innovation. We also explore the key bottlenecks limiting inference performance, how ElastixAI is tackling these issues, and why FPGAs are emerging as a compelling platform for accelerating large language model inference. ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/amelia-s-weekly-fish-fry-166630/episodes/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/amelia-s-weekly-fish-fry-166630/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.