# How Data Scientists Use Gradient Boosting for Tabular Data Page: https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-gradient-boosting-for-tabular-data Text version: https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-gradient-boosting-for-tabular-data.md Podcast: [The Data Science Podcast with Fexingo: Analytics, Machine Learning, and Data-Driven Conversations](https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831) Published: 2026-07-11T20:56:53+00:00 Episode link: https://audio.fexingo.com/business/the-data-science-podcast/episode-0104.mp3 Audio file: https://audio.fexingo.com/business/the-data-science-podcast/episode-0104.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-gradient-boosting-for-tabular-data Duration seconds: 551 ## Resource A deep dive into the enduring power of gradient boosting machines (GBMs) for structured, tabular data—the bread and butter of most real-world data science. Lucas and Luna explore why gradient boosting consistently wins Kaggle competitions and beats deep learning on many business problems. They break down the core mechanics: sequential tree-building, learning rate, and regularization. The episode focuses on a case study from a mid-size e-commerce company that used XGBoost to reduce customer churn prediction error by 18% year-over-year. They also discuss modern variants like LightGBM and CatBoost, and when to choose each. Practical guidance on hyperparameter tuning and common pitfalls (overfitting, categorical encoding) grounds the conversation in daily data-science work. Listeners will walk away understanding why gradient boosting remains a must-have in any data scientist's toolkit, especially for data with mixed data types and missing values. #GradientBoosting #XGBoost #LightGBM #CatBoost #TabularData #MachineLearning #DataScience #Kaggle #HyperparameterTuning #ChurnPrediction #EnsembleMethods #DecisionTrees #Regularization #FeatureEngineering #BusinessAnalytics #Technology #FexingoBusiness #BusinessPodcast Keep every episode free: buymeacoffee.com/fexingo ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-gradient-boosting-for-tabular-data/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-gradient-boosting-for-tabular-data.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.