# Enterprise AI Search, RAG & Agents at Scale with Vectara Page: https://stenobird.com/podcast/virtually-speaking-podcast-174853/enterprise-ai-search-rag-agents-at-scale-with-vectara Text version: https://stenobird.com/podcast/virtually-speaking-podcast-174853/enterprise-ai-search-rag-agents-at-scale-with-vectara.md Podcast: [Virtually Speaking Podcast](https://stenobird.com/podcast/virtually-speaking-podcast-174853) Published: 2026-05-18T20:26:04+00:00 Episode link: https://www.vspeakingpodcast.com/e/enterprise-ai-search-rag-agents-at-scale-with-vectara/ Audio file: https://mcdn.podbean.com/mf/web/qd38jsufqvxmrvq4/VSP-KUBECON-JEFF-V2.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/virtually-speaking-podcast-174853/episodes/enterprise-ai-search-rag-agents-at-scale-with-vectara Duration seconds: 982 ## Resource At KubeCon 2026, Jad El-Zein and Frank Denneman sit down with Jeff Chapman from Vectara to discuss how enterprise RAG, vector databases, and AI agents are evolving inside modern private AI environments. The conversation explores how Vectara integrates with VMware Private AI Foundation and VMware Cloud Foundation to help organizations scale AI applications securely across millions of documents while maintaining role-based access control, multimodal ingestion, and sovereign data protections. They also dive into enterprise search, hallucination prevention, citations, agent orchestration, long-running AI agents, GPU efficiency, and why on-prem AI infrastructure is becoming increasingly important for enterprises building production AI systems. Topics include: Enterprise RAG vs traditional search Vector databases and multimodal AI Role-based access control for AI AI agents and orchestration Sovereign AI and air-gapped environments GPU utilization and scaling AI workloads VMware Private AI Foundation integration On-prem AI economics and token costs #KubeCon #AI #PrivateAI #VMware #VCF #RAG #Agents #Kubernetes #VectorDatabase #EnterpriseAI ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/virtually-speaking-podcast-174853/episodes/enterprise-ai-search-rag-agents-at-scale-with-vectara/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/virtually-speaking-podcast-174853/enterprise-ai-search-rag-agents-at-scale-with-vectara.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.