# How Data Scientists Use Retrieval Augmented Generation for Enterprise Search Page: https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-retrieval-augmented-generation-for-enterprise-search Text version: https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-retrieval-augmented-generation-for-enterprise-search.md Podcast: [The Data Science Podcast with Fexingo: Analytics, Machine Learning, and Data-Driven Conversations](https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831) Published: 2026-07-14T20:50:27+00:00 Episode link: https://audio.fexingo.com/business/the-data-science-podcast/episode-0110.mp3 Audio file: https://audio.fexingo.com/business/the-data-science-podcast/episode-0110.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-retrieval-augmented-generation-for-enterprise-search Duration seconds: 472 ## Resource Episode 110 of The Data Science Podcast dives into Retrieval Augmented Generation (RAG) for enterprise search. Lucas and Luna explore how companies like JP Morgan and NASA are using RAG to make internal documents searchable and actionable. They discuss the key components: embedding models, vector databases like Pinecone, and large language models like GPT-4. The episode walks through a concrete example: a financial analyst querying a 10-K filing for revenue recognition policies. They cover challenges like chunking strategies, retrieval quality, and hallucination risks, plus emerging techniques like HyDE and multi-hop retrieval. By the end, listeners understand RAG's role in unlocking unstructured data at scale. #RetrievalAugmentedGeneration #EnterpriseSearch #RAG #VectorDatabases #Embeddings #LargeLanguageModels #GPT4 #Pinecone #JP Morgan #NASA #10K Filing #HyDE #MultiHopRetrieval #UnstructuredData #DataScience #Technology #FexingoBusiness #BusinessPodcast Keep every episode free: buymeacoffee.com/fexingo ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-retrieval-augmented-generation-for-enterprise-search/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-retrieval-augmented-generation-for-enterprise-search.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.