Episode

Reconfigurable Hardware: ElastixAI and The Future of Fast, Efficient AI Inference

Podcast
Amelia's Weekly Fish Fry
Published
Jun 12, 2026
Duration seconds
1143
Processing state
not_requested
Canonical source
https://techfocus.podbean.com/e/how-elastic-ai-uses-fpgas-to-crush-llm-inference-bottlenecks/
Audio
https://mcdn.podbean.com/mf/web/3mfuzdmdir43qvby/Fish_fry_-_June_12_20269fr29.mp3
JSON
/v1/public/podcasts/amelia-s-weekly-fish-fry-166630/episodes/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference
Markdown
/podcast/amelia-s-weekly-fish-fry-166630/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/amelia-s-weekly-fish-fry-166630/episodes/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/amelia-s-weekly-fish-fry-166630/reconfigurable-hardware-elastixai-and-the-future-of-fast-efficient-ai-inference.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Artificial intelligence is moving faster than ever, but as AI models continue to grow in size and complexity, the challenges surrounding inference performance are becoming impossible to ignore. In this week's podcast, ElastixAI CEO Dr. Mohammad Rastegari and I chat about how we can overcome those challenges and why a different approach to AI infrastructure is necessary for the next generation of AI innovation. We also explore the key bottlenecks limiting inference performance, how ElastixAI is tackling these issues, and why FPGAs are emerging as a compelling platform for accelerating large language model inference.