{"podcast":{"title":"The Data Science Podcast with Fexingo: Analytics, Machine Learning, and Data-Driven Conversations","slug":"the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831","podcast_index_feed_id":7871831,"rss_url":"https://feeds.fexingo.com/business/the-data-science-podcast.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/the-data-science-podcast/cover.png","author":"Fexingo","episode_count":118,"summary":"Lucas and Luna sit at a data-science workstation, two thin laptops open to scatter plots and clustering visualizations, and ask: what can we actually learn from the numbers? Each episode of The Data Science Podcast with Fexingo is a grounded, specific conversation about a single analytics problem or machine-learning method — from regularization in regression to the bias-variance trade-off in random forests. Lucas leads with a journalistic eye for how models are built and tested in the real world, citing actual case studies like how Netflix used matrix factorization for recommendations or how healthcare researchers apply survival analysis to clinical trials. Luna keeps the discussion honest, asking about data quality, feature engineering pitfalls, and whether a model’s accuracy actually translates to business value. They never resort to buzzwords: instead, they walk through the workflow from data collection to deployment, discussing trade-offs like interpretability versus performance. The show serves data scientists, analysts, and engineers who want to stay sharp on methods without the hype. Listeners walk away with a clearer understanding of why one algorithm beats another on a gi…","last_synced_at":"2026-07-19T08:17:23.323447+00:00","page_url":"https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831"},"episode":{"title":"How Data Scientists Use Data Version Control for Reproducibility","slug":"how-data-scientists-use-data-version-control-for-reproducibility","published_at":"2026-07-13T08:53:45+00:00","page_url":"https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-data-version-control-for-reproducibility","show_page_url":"https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831","url":"https://audio.fexingo.com/business/the-data-science-podcast/episode-0107.mp3","audio_url":"https://audio.fexingo.com/business/the-data-science-podcast/episode-0107.mp3","summary":"Lucas and Luna break down why data version control (DVC) has become as essential as Git for machine learning teams. They trace the problem through a concrete example: a fraud detection model at a fintech company where a missing dataset version caused a 15 percent drop in recall. The episode walks through how DVC tracks data snapshots, pipeline stages, and model artifacts—without duplicating massive files—using a simple declarative YAML config. Lucas explains the difference between DVC's approach and Git LFS, and why tools like Pachyderm and DVC solve overlapping but distinct problems. The hosts also discuss how versioning interacts with feature stores and CI/CD for ML, and why the field is moving toward treating data with the same discipline as source code. No fluff, just a focused look at one practice that separates professional data teams from the rest. #DataVersionControl #DVC #MLOps #Reproducibility #MachineLearning #DataScience #GitForData #Pachyderm #LFS #DataPipeline #FeatureStore #CI/CD #FraudDetection #Fintech #MLPipeline #DataGovernance #Technology #FexingoBusiness Keep every episode free: buymeacoffee.com/fexingo","meta_description":"Lucas and Luna break down why data version control (DVC) has become as essential as Git for machine learning teams. They trace the problem through a concr…","key_points":[],"chapters":[],"topics":[],"duration_seconds":756,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-data-version-control-for-reproducibility/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-data-version-control-for-reproducibility.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}